Comparing prices across every provider we track...
AI Endpoints · $0.31/mo
Vultr's serverless inference offering is dramatically cheaper on both base price and per-token costs, though the model is smaller and less capable than Mistral-Small. OVHcloud's hidden costs (IPv4 billing, APAC egress changes post-2026, storage retrieval fees) further erode value. Vultr is the clear winner for cost-conscious workloads that can tolerate a smaller model.
30x cheaper base price and 31x cheaper per-token costs (input and output both $0.01/M vs OVHcloud's $0.31/M output), plus 100% network uptime SLA.
Caveat: Model is 8B parameters versus Mistral's 24B, so inference quality and capability will be noticeably lower for complex tasks; verify bandwidth overage rates for your data center.
Before you switch