Mistral
Use models in a few clicks with our platform. Download our open models for deep access.
Mixtral is a powerful and fast model adaptable to many use-cases. While being 6x faster, it matches or outperform Llama 2 70B on all benchmarks, speaks many languages, has natural coding abilities. It handles 32k sequence length. You can use it through our API, or deploy it yourself (it’s Apache 2.0!).
Pricing: Per token usage
Mistral Alternatives
Explore 92 products in the Inference APIs category. View all Mistral alternatives.
Varion
OpenAI-compatible proxy that trims LLM token usage before requests reach your provider
Scaleway
European serverless AI inference APIs, 100% hosted in Europe
Alibaba Cloud Model Studio
Hosted API access to Qwen and third-party models across six global regions
Packet.ai
On-demand NVIDIA GPU cloud with per-second billing, SSH, CLI, and API access
Work on Mistral? Feature it at the top of Inference APIs.
Is your product missing?