Serverless Inference · $0.55/mo
OVHcloud's AI Endpoints offering is significantly cheaper on base pricing ($0.01/mo vs $0.55/mo), but the plans serve different use cases: OVHcloud's bge-m3 is an embedding model while Vultr's MiMo-V2.5-Pro appears to be a general inference tier. If your workload matches embedding tasks, OVHcloud offers strong value; however, token pricing and hidden costs (IPv4 billing, Windows surcharges, regional traffic changes post-2026) require careful evaluation.
Base price is 55x cheaper and input token rate matches at $0.01/1M, making it ideal if your inference needs are embedding-focused.
Caveat: Model is specialized (embeddings only, not general inference); APAC outbound traffic loses free-egress after mid-2026; IPv4 and Windows licenses add cost if needed.
Before you switch