Comparing prices across every provider we track...
AI Endpoints · $0.11/mo
Vultr's serverless inference offering is substantially cheaper at $0.01/mo with identical $0.01/M token pricing for both input and output, compared to OVHcloud's $0.11/mo base plus $0.11/1M tokens. While the model differs (Llama vs Mistral), Vultr delivers 10x lower per-token costs and includes a 100% network uptime SLA, making it a compelling alternative if model compatibility permits.
10x cheaper base price and 10x cheaper per-token rates ($0.01 vs $0.11/1M tokens) with explicit 100% network uptime SLA.
Caveat: Different model (Llama-based safety guard vs Mistral-7B-Instruct); verify functional equivalence for your use case and check bandwidth overage rates for your data center.
Before you switch