Model details
Mistral-7B-Instruct-v0.3
Built as a decoder-only transformer, this model serves as an instruction-tuned iteration of the 7.3 billion parameter base architecture. It is designed to excel in conversational environments, prioritizing a balance between speed and performance for daily utility. The architecture incorporates an extended vocabulary of 32,768 tokens and utilizes the v3 tokenizer, which collectively enable more nuanced language processing and native support for function calling. By integrating specialized tokens for tool availability and results, the model is engineered to act as a reliable agent for structured tasks.
The model benefits from a lineage of iterative improvements, building upon the foundations of its predecessors to enhance usability and instruction adherence. It supports advanced customization through parameter-efficient fine-tuning techniques such as P-Tuning and LoRA, allowing developers to adapt the model for specific workflows using frameworks like NVIDIA NeMo. With its optimized design, the model provides a robust solution for developers seeking a high-performing, compact system capable of handling complex prompts and maintaining consistent output quality across diverse text-based applications.
Quick Info
Powered by- Provider
- OVHcloud AI Endpoints
- Model key
- mistral-7b-instruct-v0.3
- Release date
- Apr 1, 2025
- Last updated
- Apr 1, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.11
- Output token cost
- $0.11
Limits
- Output tokens
- 65,536 tokens
- Context window
- 65,536 tokens