Currently listed through these providers:
Model details
Ministral 3B
Ministral 3B is the smallest member of the Ministral 3 family, engineered as a compact powerhouse for edge and on-device deployment. With 3 billion parameters, it carries a vision-enabled architecture that processes both text and images while maintaining the low-latency, compute-efficient profile demanded by resource-constrained environments. The design prioritizes high performance across diverse hardware, from local setups to embedded systems, and it targets use cases like orchestrating agentic workflows and building specialist task workers. According to benchmark comparisons, it outperforms larger models such as Mistral 7B on most evaluations, establishing a new frontier for sub-10B language models in knowledge, commonsense reasoning, and function-calling.
The model sits at the intersection of efficiency and capability, built to excel in situations where running larger models would be impractical. Its 128k to 256k context support enables it to handle long-horizon tasks and multi-step reasoning despite its compact size. As an open-weights model, it invites fine-tuning and customization for domain-specific applications, making it adaptable for developers who need a reliable backbone for agentic pipelines or specialized AI workers operating at the edge.
Quick Info
Powered by- Provider
- DevPass (LLM Gateway)
- Model key
- ministral-3b-2512
- Release date
- Dec 2, 2025
- Last updated
- Dec 2, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.10
- Output token cost
- $0.10
Limits
- Output tokens
- 8,192 tokens
- Context window
- 131,072 tokens
Transparent token rates
Compare mistral pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Ministral 3B
No articles yet. Fetch the latest news to show it here.