Currently listed through these providers:
Model details
Mistral Nemo 12B Instruct
Mistral NeMo 12B Instruct is a 12-billion-parameter language model positioned as a compact, efficient option for developers and researchers who need multilingual understanding and generation. The Together AI listing frames it as an advanced model aimed at reasoning, code, and multilingual tasks, making it well suited to applications such as chat assistants, retrieval-augmented workflows, and lightweight code helpers where smaller footprint matters. Its design intent emphasizes breadth across languages and tasks while keeping resource demands modest compared to larger frontier models.
The model is distributed as an NVIDIA NIM container and is reachable through Together AI under the nv-mistralai routing path, with NVIDIA also hosting the corresponding container in its NGC catalog. To use the model on Together AI, it must first be deployed on a Dedicated Endpoint rather than called as a standard hosted API, which shapes how teams integrate it into production pipelines. This NIM-based packaging points to a focus on streamlined enterprise deployment, giving teams a managed route to run a multilingual instruct model with consistent inference behavior across environments.
Quick Info
Powered by- Provider
- Inference
- Model key
- mistral/mistral-nemo-12b-instruct
- Release date
- Jan 1, 2025
- Last updated
- Jan 1, 2025
- Knowledge cutoff
- 2024-12
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.038
- Output token cost
- $0.10
Limits
- Output tokens
- 4,096 tokens
- Context window
- 16,000 tokens
Latest news about Mistral Nemo 12B Instruct
No articles yet. Fetch the latest news to show it here.