Comparing prices across every provider we track...
AI Endpoints · $0.09/mo
Vultr's Llama 3.1 offering is significantly cheaper on both base price and per-token costs, though it uses a different model. OVHcloud's gpt-oss-120b is a larger, more capable model, so the choice depends on whether you need that extra capacity or can accept a smaller alternative at roughly 90% lower cost.
Input and output tokens both cost $0.01/M versus OVHcloud's $0.09/M input, delivering 9x cheaper inference at the same base tier price.
Caveat: Model is 8B parameters versus OVHcloud's 120B, so accuracy and reasoning capability will be noticeably lower for complex tasks.
Before you switch