Currently listed through these providers:
Model details
Llama-3.3-70B-Instruct
Llama-3.3-70B-Instruct is a Meta-released large language model offered through the Abacus hosted routing layer. AWS Bedrock publishes an official model documentation page for Llama 3.3 70B Instruct, confirming Meta-authorized availability of this instruction-tuned variant on a major hyperscaler. NVIDIA's NGC catalog also lists the model under Meta's organization namespace, with associated NeMo and Triton runtime configurations that allow enterprise teams to deploy it on optimized GPU stacks alongside other open Llama-family offerings.
Because the model is distributed as an instruction-tuned open-weights checkpoint within the Llama 3.3 family, it fits workflows where teams want a capable general-purpose chat and assistant model they can self-host or access through multiple clouds without vendor lock-in. Its presence on both AWS Bedrock and NVIDIA NGC means practitioners can choose between managed API access and self-managed GPU deployment paths. Teams adopting it should rely on Abacus for routing and on Meta's primary release channels for the authoritative model card and benchmark results.
Quick Info
Powered by- Provider
- Abacus
- Model key
- meta-llama/Meta-Llama-3.3-70B-Instruct
- Release date
- Dec 6, 2024
- Last updated
- Dec 6, 2024
- Knowledge cutoff
- 2023-12
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.59
- Output token cost
- $0.79
Limits
- Output tokens
- 8,192 tokens
- Context window
- 131,072 tokens
Latest news about Llama-3.3-70B-Instruct
Videos about Llama-3.3-70B-Instruct
More models around Llama-3.3-70B-Instruct
This exact model name is also listed by 24 other providers.