Currently listed through:
Model details
Mistral Nemo
Mistral Nemo is a 12 billion parameter language model that represents a deliberate step toward efficiency and accessibility, developed by Mistral AI in collaboration with NVIDIA. Its most distinctive technical feature is the Tekken tokenizer, trained across more than one hundred languages, which delivers roughly 30% better compression for source code compared to earlier Mistral tokenizers, alongside two times better compression for Korean and three times for Arabic. This combination of compression gains and a 128k token context window means the model can process long documents and codebases while consuming fewer tokens per task, directly translating to lower operational costs. The model supports eleven languages including English, French, German, Spanish, Italian, Portuguese, Chinese, Japanese, Korean, Arabic, and Hindi, positioning it as a practical multilingual workhorse rather than a narrowly focused English-only model.
The collaboration between Mistral and NVIDIA brought quantization-aware training into the picture, enabling FP8 inference without performance degradation and giving the model a deployment efficiency edge. Mistral Nemo ships under the Apache 2.0 license with both base and instruct weights available on HuggingFace, making it accessible for fine-tuning and self-hosting. The model was designed as a drop-in replacement for Mistral 7B, aiming to deliver enhanced instruction following, stronger multi-turn conversation quality, and improved code generation while keeping the same deployment footprint. These attributes make Mistral Nemo well suited for developers and teams looking for an open, capable, and cost-efficient foundation model that performs across languages and coding tasks.
Quick Info
Powered by- Provider
- Kilo Gateway
- Model key
- mistralai/mistral-nemo
- Release date
- Jul 1, 2024
- Last updated
- Jul 1, 2024
- Knowledge cutoff
- 2024-07
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.019
- Output token cost
- $0.03
Limits
- Output tokens
- 16,384 tokens
- Context window
- 131,072 tokens
Transparent token rates
Compare mistral-nemo pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Mistral Nemo
No articles yet. Fetch the latest news to show it here.
Videos about Mistral Nemo
More models around Mistral Nemo
This exact model name is also listed by 7 other providers.