Currently listed through these providers:
Model details
Open Mistral Nemo
Open Mistral Nemo is a chat-mode language model positioned for cost-sensitive automation and background workloads where scale matters more than frontier capability. Aggregator listings describe it as well-suited to high-volume processing, extended document handling, and agentic pipelines that benefit from the cataloged API limit token context window combined with tool choice and structured-output support. The model's flat, symmetric pricing is unusually economical for the Mistral lineup, making it a practical choice for batch summarization, routine extraction, and continuous back-office tasks where latency and reasoning depth are secondary concerns.
Because Open Mistral Nemo carries a listed deprecation date in third-party catalogs, teams adopting it should plan a migration path toward successor Mistral chat models while taking advantage of its remaining strengths. Its confirmed capabilities—temperature control within a 0–1.5 range, tool calling, response-schema support, and memory support—cover the core needs of typical orchestration frameworks, while explicitly unsupported areas such as reasoning effort controls, verbosity tuning, thinking levels, computer use, and deep research signal that more ambitious workflows belong on newer releases. For now, the model fits a narrow but well-defined role: dependable, inexpensive text generation inside long-context agent loops.
Quick Info
Powered by- Provider
- Mistral
- Model key
- open-mistral-nemo
- Release date
- Jul 1, 2024
- Last updated
- Jul 1, 2024
- Knowledge cutoff
- 2024-07
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.15
- Output token cost
- $0.15
Limits
- Output tokens
- 128,000 tokens
- Context window
- 128,000 tokens