Currently listed through these providers:
Model details
Llama 3.3 70B Instruct
Llama 3.3 70B Instruct is a large language model in Meta's Llama family that uses a transformer architecture to interpret natural-language prompts and produce coherent, contextually relevant text. Described as an instruct-tuned model, it is geared toward following directions and completing tasks expressed in plain language, making it well suited for chatbots, content creation tools, and language translation services where instruction following and conversational fluency matter most.
The model is offered through multiple inference channels, including an FP8 inference listing on Lambda and a dedicated model-card page on Amazon Bedrock, signaling broad third-party availability for production deployments. Its flexible design lets it adapt to a wide range of topics and queries, and pairing the FP8-quantized variant with managed-cloud hosting gives integrators a practical path to balance cost, throughput, and response quality when embedding it in real-world applications.
Quick Info
Powered by- Provider
- Merge Gateway
- Model key
- meta/llama-3.3-70b-instruct
- Release date
- Dec 6, 2024
- Last updated
- Dec 6, 2024
- Knowledge cutoff
- 2023-12
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.22
- Output token cost
- $0.50
Limits
- Output tokens
- 32,768 tokens
- Context window
- 131,072 tokens
Latest news about Llama 3.3 70B Instruct
Videos about Llama 3.3 70B Instruct
More models around Llama 3.3 70B Instruct
This exact model name is also listed by 24 other providers.