AI Endpoints · $0.18/mo
Vultr's serverless inference offering is dramatically cheaper on both base price and per-token costs, though the model is smaller (8B vs 20B parameters). OVHcloud's hidden costs—IPv4 charges, potential egress billing changes in APAC after mid-2026, and storage retrieval fees—further erode its value proposition compared to Vultr's simpler, more transparent pricing.
18x cheaper monthly base price and 18x cheaper per-token costs (input and output both $0.01/M vs OVHcloud's $0.18/M output), with 100% network uptime SLA and no mention of egress billing changes.
Caveat: Model is significantly smaller (8B vs 20B parameters), so may not handle complex tasks; verify model capability matches your use case before migrating.
Before you switch