Model details
Llama 3.3 70b Instruct
Llama 3.3 70B Instruct carries forward Meta's open-weight strategy with a focus on multilingual conversational applications. This 70 billion parameter model is pretrained and instruction-tuned, designed to excel in dialogue use cases where fluency across multiple languages matters. The model supports eight languages including English, German, French, Italian, Portuguese, Hindi, Spanish, and Thai, with a generous 131,072 token context window that enables extended multi-turn conversations and complex reasoning across language boundaries. Its architecture builds on the proven Llama foundation while adding optimizations specifically for instruction-following and conversational coherence.
Benchmark evidence shows the model delivers meaningful improvements over its predecessors, with Oracle's documentation noting it surpasses both Llama 3.1 70B and Llama 3.2 90B on text-based tasks. The instruction-tuned variant is available for fine-tuning, giving enterprises and developers a pathway to customize the base model for domain-specific applications. FlashAttention-4 support enables efficient inference on modern hardware, making the model practical for production deployments. Together AI lists it among models excelling in dialogue scenarios, positioning it as a strong choice for applications requiring natural multilingual conversation at scale.
Quick Info
Powered by- Provider
- Nvidia
- Model key
- meta/llama-3.3-70b-instruct
- Release date
- Nov 26, 2024
- Last updated
- Nov 26, 2024
- Input modalities
- Output modalities
- Capabilities
Cost
A provider subscription or plan supersedes token-based pricing for this model.
Limits
- Output tokens
- 4,096 tokens
- Context window
- 128,000 tokens
Latest news about Llama 3.3 70b Instruct
No articles yet. Fetch the latest news to show it here.