Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Tempr Gateway logo

Model details

Mistral Small (latest)

Mistral Small is the compact tier in Mistral AI's lineup, positioned for developers who want a small, fast model that can run on local hardware while still covering a useful slice of general-purpose tasks. The family lineage is anchored by Mistral Small 3, which Mistral AI described as a latency-optimized 24-billion-parameter model released under the Apache 2.0 license, signaling a shift from earlier Mistral Research License terms toward permissive open weights for general-purpose releases.

In practical terms, Mistral Small is aimed at workloads where quick responses and the ability to self-host matter more than top-end reasoning, such as chat assistants, on-device copilots, retrieval-augmented agents, and prototyping against larger proprietary systems. Mistral AI positioned Small 3 as competitive with much larger models like Llama 3.3 70B and Qwen 32B while running more than three times faster on the same hardware, and offered it as an open replacement for opaque smaller proprietary models, a framing that continues to shape how the Small line fits into mixed-model deployments today.

Tempr Gatewaymistral/mistral-small-latestmistral-small

Quick Info

Powered by
Provider
Tempr Gateway
Model key
mistral/mistral-small-latest
Release date
Mar 16, 2026
Last updated
Mar 16, 2026
Knowledge cutoff
2025-06
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.15
Output token cost
$0.60

Limits

Output tokens
256,000 tokens
Context window
256,000 tokens

Transparent token rates

Compare Mistral Small (latest) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Mistral Small (latest)

No articles yet. Fetch the latest news to show it here.

Videos about Mistral Small (latest)

More models around Mistral Small (latest)