Comparing prices across every provider we track...
Serverless Inference · $0.3/mo
OVHcloud's AI Endpoints offering is significantly cheaper on base pricing ($0.01/mo vs $0.3/mo) and token costs ($0.01/1M input vs $0.30/M), but the product appears to be an embedding model rather than a general-purpose inference endpoint. Verify that OVHcloud's bge-m3 matches your actual inference workload before switching, as the lower cost may reflect a narrower use case.
Dramatically lower base fee and input token pricing makes this attractive if embedding-only inference meets your needs.
Caveat: bge-m3 is a specialized embedding model, not a general-purpose LLM inference service; output token pricing not listed, suggesting this may not support text generation.
Before you switch