Comparing prices across every provider we track...
Serverless Inference · $0.3/mo
Vultr's Kimi-K2.6 is competitively priced for inference, but OVHcloud's bge-m3 embedding model offers significantly lower token costs if your use case matches embedding workloads. However, the two products serve different purposes: Vultr targets general-purpose LLM inference while OVHcloud specializes in embeddings, so direct comparison is limited.
Dramatically lower base fee and input token cost ($0.01/1M vs $0.30/M), ideal if you need embeddings rather than full LLM inference.
Caveat: This is an embedding model, not a general-purpose LLM like Kimi-K2.6; output token pricing not listed, suggesting it may not apply to this product type.
Before you switch