Currently listed through these providers:
Model details
Mistral Nemo Instruct 2407 TEE
Mistral Nemo Instruct 2407 TEE is the Chutes-hosted endpoint of the Mistral–NVIDIA collaboration known for efficient multilingual chat and local-friendly deployment. The underlying Mistral Nemo line was designed as a compact, open-weights instruction-tuned model that balances conversational quality with the ability to run on modest hardware, and the unsloth-branded TEE variant on Chutes carries that same lineage. Because it is distributed as an open-weight checkpoint, teams that need transparency into the model file or want to replicate behavior can audit and self-host the weights, while still taking advantage of an inexpensive hosted inference path through Chutes for production traffic.
Practically, this endpoint is a good fit for straightforward chat workloads where a very large context window, generous headroom for long completions, and low per-token costs matter more than advanced agent-style features. On Chutes the published capability set is limited to temperature control, with tool calling, reasoning, and structured output not exposed for this specific route, so it is best suited for natural-language generation, summarization, drafting, and multilingual Q&A rather than tool-mediated pipelines. Teams that already lean on the broader Mistral Nemo 2407 release will find a familiar instruction-following profile here, and the large context ceiling makes it practical for long-document analysis, code review, or extended conversation threads at a budget-friendly price point.
Quick Info
Powered by- Provider
- Chutes
- Model key
- unsloth/Mistral-Nemo-Instruct-2407-TEE
- Release date
- Jul 1, 2024
- Last updated
- Jul 1, 2024
- Knowledge cutoff
- 2024-07
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.0245
- Output token cost
- $0.0978
Limits
- Output tokens
- 131,072 tokens
- Context window
- 131,072 tokens