Currently listed through these providers:
Model details
Llama 3.3 70B
Llama 3.3 70B is Meta's instruction-tuned 70-billion-parameter text model designed for multilingual dialogue and general text tasks. It belongs to the Llama family of large language models and processes text input to generate text output, supporting a 128,000-token context window that lets it reason over long documents, sustained conversations, and multi-step instructions within a single interaction. The model's lineage as a Meta-released, openly available weights checkpoint makes it a practical choice for teams that want strong general-purpose language understanding without depending on a closed proprietary system.
For practitioners, Llama 3.3 70B is positioned as a step up in quality compared with earlier Llama generations, with Meta claiming it delivers better performance than the larger Llama 3.2 90B and the previous-generation Llama 3.1 70B on a range of text tasks. That combination of a relatively compact 70B footprint and the cataloged API limit context window makes it well suited to assistants, retrieval-augmented generation, summarization, code-adjacent reasoning, and multilingual customer support scenarios where a single open-weights model needs to balance depth of reasoning with deployment flexibility across hosted platforms.
Quick Info
Powered by- Provider
- NanoGPT
- Model key
- TEE/llama3-3-70b
- Release date
- Jul 3, 2025
- Last updated
- Jul 3, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.75
- Output token cost
- $2.75
Limits
- Input tokens
- 128,000 tokens
- Output tokens
- 16,384 tokens
- Context window
- 128,000 tokens
Latest news about Llama 3.3 70B
No articles yet. Fetch the latest news to show it here.