Currently listed through these providers:
Model details
Mistral: Mixtral 8x22B Instruct
Mixtral 8x22B Instruct is built on a sparse Mixture-of-Experts architecture that leverages a total of 141B parameters while activating only 39B per forward pass. This design intent allows the model to achieve high-level performance in complex reasoning, mathematics, and coding tasks while maintaining faster inference speeds than dense models of a similar scale. By utilizing this efficient expert-based structure, the model is engineered to handle demanding workloads, making it a robust choice for applications that require both depth of knowledge and operational efficiency.
As an instruction-tuned variant, this model is specifically optimized for multi-turn dialogue and following intricate user prompts. Its training lineage emphasizes versatility, resulting in strong fluency across English, French, Italian, German, and Spanish. The model includes native function calling and constrained output capabilities, which are essential for building agentic workflows and structured data applications. With its ability to perform well on benchmarks like GSM8K and MATH, it serves as a reliable foundation for developers seeking a balance between large-scale reasoning power and practical, real-world utility.
Quick Info
Powered by- Provider
- Kilo Gateway
- Model key
- mistralai/mixtral-8x22b-instruct
- Release date
- Apr 17, 2024
- Last updated
- Apr 17, 2024
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $2.00
- Output token cost
- $6.00
Limits
- Output tokens
- 52,428 tokens
- Context window
- 65,536 tokens
Transparent token rates
Compare mistral pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.