Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Vercel AI Gateway logo

Model details

Ministral 3B (latest)

Ministral 3B was introduced by Mistral AI as part of the Ministral family of edge-focused models, designed to bring capable language understanding to on-device and low-latency environments. Alongside the larger Ministral 8B, it targets scenarios where privacy-preserving local inference matters, including offline assistants, smart-device translation, local analytics, and robotics control loops. The model is positioned as a compute-efficient option that can also serve as a fast intermediary in multi-step workflows, handling input parsing, task routing, and API invocation when paired with larger models. Its compact footprint makes it well suited to specialist task workers that need to be fine-tuned for narrow roles rather than general-purpose reasoning.

Architecturally, Ministral 3B is a sub-10 billion parameter model optimized for efficiency in the small-model category, and the broader Ministral family supports extended context handling that enables longer documents and richer conversational state than typical edge models. Mistral highlights strong performance across knowledge, commonsense reasoning, and function calling for its size, making it attractive for agentic pipelines where a lightweight model must reliably call external tools. Open-weight availability further extends its practical appeal, allowing teams to self-host, fine-tune, or distill the model into specialized variants while keeping inference costs and latency low. For builders, the combination of open weights, tool-use support, and a small parameter count makes Ministral 3B a flexible building block for privacy-sensitive and latency-critical applications.

Vercel AI Gatewaymistral/ministral-3bministral

Quick Info

Powered by
Provider
Vercel AI Gateway
Model key
mistral/ministral-3b
Release date
Oct 1, 2024
Last updated
Oct 4, 2024
Knowledge cutoff
2024-10
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.04
Output token cost
$0.04

Limits

Output tokens
128,000 tokens
Context window
128,000 tokens

Transparent token rates

Compare Ministral 3B (latest) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Ministral 3B (latest)

Vercel AI Gateway

CoverageBenchmark

The OpenOCR benchmark page (published 2026-08-22) provides direct evidence for the Ministral 3B variant "ministral-3b-2512," measured on Mistral's own provider offering under the identifier "mistral-ministral-3b-latest." The evaluation covers 11 documents from the 'ocr-shared-11-normalized/v2' corpus, scoring the model Per-document results include a grocery receipt at 100.0% (1.4s), a café receipt at 86.7% (1.0s), a handwritten letter at 100.0% (791ms), and a handwritten notes sample at 100.0% (700ms), with measurements dated 2026-08-07. The page attributes the model to Mistral and explicitly names the "ministral-3b-2512" variant alo

Vercel AI Gateway

CoverageBenchmark

Compare Claude Opus 4.5 vs Min istral 3 (3B Reasoning 2512): input $5/M vs $0.1/M, output $25/M vs $0.1/M tokens. Min istral 3 (3B Reasoning 2512) is 14900% cheaper overall. Full API cost breakdown, context window, and benchmark comparison.

Vercel AI Gateway

CoverageBenchmark

The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.

Videos about Ministral 3B (latest)

More models around Ministral 3B (latest)