Comparing prices across every provider we track...
AI Endpoints · $0.14/mo
Vultr's Llama-3.1-Nemotron offering is substantially cheaper at $0.01/M tokens (vs OVHcloud's $0.14/M), representing a 93% cost reduction on inference. While the models differ, Vultr delivers better value for comparable serverless inference workloads, though you should confirm model performance parity for your use case.
Input and output tokens both priced at $0.01/M, a 14x reduction versus OVHcloud's $0.14/M rate.
Caveat: Model is Llama-based (not Mistral), so verify inference quality and latency match your requirements; bandwidth overages are location-dependent and not fully transparent.
Before you switch