Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Vercel AI Gateway logo

Model details

Mistral Small (latest)

Mistral Small 4 is a sparse mixture-of-experts model designed to deliver strong performance without proportional compute cost. With 119 billion total parameters that activate only 6 billion per token during inference, it balances scale and efficiency in a way that supports fast text responses alongside logical reasoning and image understanding in a single multimodal package. This architecture places it in an interesting position between compact models that sacrifice capability for speed and larger dense models that demand more infrastructure. The hybrid design unifying instruct, reasoning, and coding capabilities suggests it was built to handle the breadth of everyday AI tasks rather than excelling at a single narrow benchmark.

The model arrives as the latest iteration in Mistral's Small family, building on the momentum of Mistral Small 3's Apache 2.0 release. Early variants in this line prioritized latency and efficiency, achieving competitive results against much larger models while maintaining throughput suitable for high-volume workloads. Mistral Small 4 inherits this lineage while adding multimodal support and expanded context handling for longer documents and conversations. The Apache licensing carried forward from earlier releases makes it practical for enterprises to deploy, customize, and run at scale without licensing friction. This combination of architectural efficiency, broad capability coverage, and open deployment options positions it as a practical choice for cost-sensitive automations and background processing where reliability and speed matter more than frontier-level scale.

Vercel AI Gatewaymistral/mistral-smallmistral-small

Quick Info

Powered by
Provider
Vercel AI Gateway
Model key
mistral/mistral-small
Release date
Sep 17, 2024
Last updated
Mar 16, 2026
Knowledge cutoff
2025-06
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.15
Output token cost
$0.60

Limits

Output tokens
4,000 tokens
Context window
262,144 tokens

Transparent token rates

Compare Mistral Small (latest) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Mistral Small (latest)

Vercel AI Gateway

Coverage

Mistral AI has released Mistral Small 4, combining fast text responses, logical reasoning, and image processing in one model.

Videos about Mistral Small (latest)

More models around Mistral Small (latest)