Currently listed through:
Model details
Mixtral 8x22B
Mixtral 8x22B utilizes a sparse mixture-of-experts architecture, comprising eight distinct expert submodels that total 141 billion parameters. By activating only 39 billion parameters during inference, the model achieves a balance between high-level reasoning capabilities and computational efficiency. This design allows it to outperform many dense models of similar or larger scale, making it a versatile choice for complex tasks such as mathematical reasoning, code generation, and nuanced multilingual communication in English, French, Italian, German, and Spanish.
Built as a natural evolution of the mixtral family, this model is engineered to provide a superior performance-to-cost ratio for developers and researchers. Its native support for function calling and constrained output modes makes it well-suited for modernizing technical stacks and building agentic applications. By offering a large context window, the model excels at precise information recall across extensive documents, serving as a robust foundation for fine-tuning and specialized deployments where both speed and depth of knowledge are required.
Quick Info
Powered by- Provider
- Mistral
- Model key
- open-mixtral-8x22b
- Release date
- Apr 17, 2024
- Last updated
- Apr 17, 2024
- Knowledge cutoff
- 2024-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $2.00
- Output token cost
- $6.00
Limits
- Output tokens
- 64,000 tokens
- Context window
- 64,000 tokens
Latest news about Mixtral 8x22B
No articles yet. Fetch the latest news to show it here.