Currently listed through these providers:
Model details
Ministral 3B
Ministral 3B is a state-of-the-art small language model engineered to balance high-level performance with the efficiency required for edge computing and on-device applications. Designed to excel in the sub-10B parameter category, the model focuses on delivering robust knowledge, commonsense reasoning, and precise function-calling capabilities. Its architecture is specifically optimized for low-latency inference, making it a versatile tool for developers building specialized task workers or orchestrating multi-step agentic workflows that demand both speed and reliability.
Built to serve as an efficient intermediary in complex AI systems, the model can be paired with larger language models to streamline function-calling tasks. Its design lineage emphasizes adaptability, allowing for fine-tuning to meet specific application needs. By providing strong performance in benchmarks against other compact models, it offers a practical solution for developers who need to deploy intelligent, responsive AI across diverse hardware environments, including local setups, without sacrificing the reasoning depth required for modern generative AI tasks.
Quick Info
Powered by- Provider
- Azure Cognitive Services
- Model key
- ministral-3b
- Release date
- Oct 22, 2024
- Last updated
- Oct 22, 2024
- Knowledge cutoff
- 2024-03
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.04
- Output token cost
- $0.04
Limits
- Output tokens
- 8,192 tokens
- Context window
- 128,000 tokens
Transparent token rates
Compare Ministral 3B pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Ministral 3B
No articles yet. Fetch the latest news to show it here.
Videos about Ministral 3B
More models around Ministral 3B
This exact model name is also listed by 5 other providers.