Comparing prices across every provider we track...
AI Endpoints · $0.31/mo
Vultr's serverless inference offering is dramatically cheaper on both base price and per-token costs, though the model is smaller and less capable. OVHcloud's Mistral-Small is a more powerful model, but you'll pay significantly more unless your use case doesn't require the extra performance.
Token pricing is 31x cheaper ($0.01/M vs $0.31/M for output), with a 100% network uptime SLA and no regional egress restrictions.
Caveat: The 8B model is significantly smaller and less capable than Mistral-Small-3.2-24B; verify it meets your inference quality requirements before migrating.
Before you switch