Comparing prices across every provider we track...
AI Endpoints · $1.01/mo
Vultr's Llama-3.1-Nemotron model is 99% cheaper on base pricing and token costs, though it is a different model class (safety-guard vs vision-language). OVHcloud's Qwen2.5-VL-72B is a specialized vision-language model; if that capability is essential, OVHcloud is your only option here, but verify whether you actually need vision features before paying the premium.
Identical $0.01/M token pricing for both input and output with negligible base cost, plus 100% network uptime SLA.
Caveat: Model is 8B safety-guard, not a 72B vision-language model; only comparable if you do not need vision capabilities.
Before you switch