Currently listed through these providers:
Model details
Mistral Small 4 119B Thinking
Mistral Small 4 119B Thinking is the dedicated reasoning variant in Mistral's small-model family, tuned for step-by-step analysis while keeping the footprint modest enough for cost-sensitive deployments. On the NanoGPT route it is exposed with native reasoning output, tool calling, and structured output support, which makes it a practical fit for agentic pipelines, multi-step research tasks, and retrieval-augmented workflows where the model needs to plan, invoke external tools, and return machine-readable results. Its multimodal input handling also broadens the use cases to vision-grounded tasks such as document understanding or chart reasoning, with text as the output modality.
As an open-weight release, the Thinking variant can be self-hosted or routed through compatible providers, giving teams flexibility between managed inference and on-premises control. The NanoGPT endpoint pairs the model with a 262,144-token context window and a 16,384-token output ceiling, well suited to long-document analysis and extended chain-of-thought traces. Priced at $0.40 per million input tokens and $1.40 per million output tokens, it sits in the mid-tier of reasoning models, targeting developers who want stronger deliberation than base small models without the cost of flagship reasoning systems.
Quick Info
Powered by- Provider
- NanoGPT
- Model key
- mistralai/mistral-small-4-119b-2603:thinking
- Release date
- Mar 17, 2026
- Last updated
- Mar 17, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.40
- Output token cost
- $1.40
Limits
- Input tokens
- 262,144 tokens
- Output tokens
- 16,384 tokens
- Context window
- 262,144 tokens