Sulat.com
AI models
Kilo Gateway logo

Model details

Mistral Nemo

Mistral Nemo is a 12 billion parameter language model that represents a deliberate step toward efficiency and accessibility, developed by Mistral AI in collaboration with NVIDIA. Its most distinctive technical feature is the Tekken tokenizer, trained across more than one hundred languages, which delivers roughly 30% better compression for source code compared to earlier Mistral tokenizers, alongside two times better compression for Korean and three times for Arabic. This combination of compression gains and a 128k token context window means the model can process long documents and codebases while consuming fewer tokens per task, directly translating to lower operational costs. The model supports eleven languages including English, French, German, Spanish, Italian, Portuguese, Chinese, Japanese, Korean, Arabic, and Hindi, positioning it as a practical multilingual workhorse rather than a narrowly focused English-only model.

The collaboration between Mistral and NVIDIA brought quantization-aware training into the picture, enabling FP8 inference without performance degradation and giving the model a deployment efficiency edge. Mistral Nemo ships under the Apache 2.0 license with both base and instruct weights available on HuggingFace, making it accessible for fine-tuning and self-hosting. The model was designed as a drop-in replacement for Mistral 7B, aiming to deliver enhanced instruction following, stronger multi-turn conversation quality, and improved code generation while keeping the same deployment footprint. These attributes make Mistral Nemo well suited for developers and teams looking for an open, capable, and cost-efficient foundation model that performs across languages and coding tasks.

Kilo Gatewaymistralai/mistral-nemomistral-nemo

Quick Info

Powered by
Provider
Kilo Gateway
Model key
mistralai/mistral-nemo
Release date
Jul 1, 2024
Last updated
Jul 1, 2024
Knowledge cutoff
2024-07
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.019
Output token cost
$0.03

Limits

Output tokens
16,384 tokens
Context window
131,072 tokens

Transparent token rates

Compare mistral-nemo pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Mistral Nemo

No articles yet. Fetch the latest news to show it here.

Videos about Mistral Nemo

More models around Mistral Nemo