Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Together AI logo

Model details

Llama 3.3 70B

Llama 3.3 70B is a 70-billion parameter multilingual large language model designed for high-quality dialogue generation. Built as an instruction-tuned variant, it brings together the strengths of open-weight accessibility with competitive performance against both open-source and closed commercial chat models. The model supports eight languages including English, German, French, Italian, Portuguese, Hindi, Spanish, and Thai, making it well-suited for applications requiring multilingual communication. Its optimized architecture delivers strong results in a more compact form factor than previous flagship models, bringing advanced AI capabilities within reach of developers and organizations with standard computational resources.

Building on the Llama lineage, this model demonstrates meaningful improvements over its predecessors, outperforming both Llama 3.1 70B and Llama 3.2 90B on text tasks, while achieving performance comparable to the larger Llama 3.1 405B. Benchmark results show particular strength in reasoning, mathematical problem-solving, and instruction following. The combination of open weights, fine-tuning support, and FlashAttention-4 optimization makes it practical for both rapid deployment and customization. Organizations can leverage this model for everything from conversational AI to specialized applications requiring instruction adherence and multilingual dialogue capabilities.

Together AImeta-llama/Llama-3.3-70B-Instruct-Turbollama

Quick Info

Powered by
Provider
Together AI
Model key
meta-llama/Llama-3.3-70B-Instruct-Turbo
Release date
Dec 6, 2024
Last updated
Jul 2, 2026
Knowledge cutoff
2023-12
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.04
Output token cost
$1.04

Limits

Output tokens
131,072 tokens
Context window
131,072 tokens

Transparent token rates

Compare Llama 3.3 70B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Llama 3.3 70B

No articles yet. Fetch the latest news to show it here.

Videos about Llama 3.3 70B

More models around Llama 3.3 70B