Comparing prices across every provider we track...
AI Endpoints · $0.74/mo
Vultr's serverless inference offering is dramatically cheaper at $0.01/M tokens (vs OVHcloud's $0.74/M), though it uses a smaller model (8B vs 70B). For workloads that don't require a 70B-parameter model, Vultr delivers 98% lower token costs. OVHcloud's 70B option is best if you specifically need that model size and can absorb the hidden costs (IPv4 charges, APAC egress changes post-2026).
Token pricing is 74x cheaper ($0.01 vs $0.74 per million tokens), with 100% network uptime SLA backing reliability.
Caveat: Model is 8B parameters, not 70B, so unsuitable if you need the larger model's reasoning capability; bandwidth overages vary by data center and are not free.
Before you switch