Sulat.com
AI models
Berget.AI logo

Model details

Llama 3.3 70B Instruct

Meta's Llama 3.3 70B Instruct is positioned as an instruction-tuned 70-billion-parameter text model aimed squarely at multilingual dialogue and general assistant-style interactions. Within the Llama family lineage, it sits alongside variants such as Llama-3.1-Nemotron-70B-Instruct on NVIDIA's NeMo NGC catalog, reflecting the broader Meta open-weights Llama series that has steadily expanded parameter scale and instruction tuning across generations. The model is built for chat and task-oriented workflows where a single general-purpose LLM needs to handle conversational reasoning, summarization, drafting, and tool-mediated answers across many languages, rather than narrow single-task use cases.

In practical terms, the model is best matched to deployments that want a capable open-weights LLM without resorting to much larger frontier weights, benefiting from Llama-lineage tuning improvements while keeping serving costs lower than flagship-scale alternatives. Because the weights are openly published, teams can self-host, fine-tune, or distill from this checkpoint for domain-specific assistants, customer support bots, internal copilots, and multilingual content workflows, and it is also offered through hosted inference providers for teams that prefer managed endpoints. It fits well as a workhorse generalist for production chat and instruction-following, while more demanding reasoning-heavy workloads may warrant routing to larger or specialized models in the wider ecosystem.

Berget.AImeta-llama/Llama-3.3-70B-Instructllama

Quick Info

Powered by
Provider
Berget.AI
Model key
meta-llama/Llama-3.3-70B-Instruct
Release date
Apr 27, 2025
Last updated
Apr 27, 2025
Knowledge cutoff
2023-12
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.99
Output token cost
$0.99

Limits

Output tokens
8,192 tokens
Context window
128,000 tokens

Transparent token rates

Compare llama pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Llama 3.3 70B Instruct

No articles yet. Fetch the latest news to show it here.

Videos about Llama 3.3 70B Instruct

More models around Llama 3.3 70B Instruct