NovitaAI
Meta has released a new model, Llama 3.3 70B Instruct, now available in GitHub Models. It provides similar performance to Llama 3.1 405B, but at a...
Model details
Llama 3.3 70B Instruct is an auto-regressive language model built on an optimized transformer architecture, designed to excel in conversational and instructional tasks. As a successor to previous iterations, it is engineered to provide high-level performance in reasoning, mathematics, and code generation while maintaining a focus on multilingual text support. Its design intent centers on flexibility and adaptability, making it a robust choice for developers building complex AI features such as chatbots, content creation tools, and language translation services that require coherent, contextually relevant responses.
The model benefits from a rigorous training lineage that utilizes a diverse mix of publicly available online data. To ensure it aligns with human preferences for helpfulness and safety, the model undergoes supervised fine-tuning and reinforcement learning with human feedback. These post-training methods allow it to achieve industry-leading results in instruction following, often matching the performance of much larger models while offering a more accessible and cost-effective option for production environments. Its architecture is further optimized for modern hardware, ensuring efficient inference for demanding, real-world applications.
Transparent token rates
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
NovitaAI
Meta has released a new model, Llama 3.3 70B Instruct, now available in GitHub Models. It provides similar performance to Llama 3.1 405B, but at a...
NovitaAI
IBM announces the availability of the updated Llama 3.3 70B Instruct model on IBM watsonx.ai™.
NovitaAI
Llama 3.3 70B Instruct is the December update of Llama 3.1 70B. The model improves upon Llama 3.1 70B (released July 2024) with advances in tool calling, multilingual text support, math and coding. The model achieves industry leading results in reasoning, math and instruction following and provides similar performance
NovitaAI
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). $0 per million input tokens, $0 per million output tokens. 65,536 token context window. Higher uptime with 14 providers. Includes independent benchmarks from Artificial Analysis.
This exact model name is also listed by 24 other providers.