Serverless Inference · $0.13/mo
Vultr's Nemotron-3-Nano pricing is reasonable for a reasoning model, but the comparison is complicated by OVHcloud offering a much cheaper embedding model (bge-m3) that solves a different use case. If you need reasoning capabilities specifically, Vultr is competitive; if embeddings suffice, OVHcloud is significantly cheaper but verify token pricing and model suitability for your workload.
Dramatically lower base and token costs ($0.01/1M input tokens vs Vultr's $0.13/M), but this is an embedding model, not a reasoning model.
Caveat: bge-m3 is not a reasoning model and cannot replace Nemotron-3-Nano-Omni for LLM inference; APAC egress loses free status after mid-2026; IPv4 and Windows license surcharges apply if needed.
Before you switch