Llama
The latest AI models from Meta, Llama-4-Scout-17B-16E-Instruct and Llama-4-Maverick-17B-128E-Instruct-FP8, are now available on GitHub Models.
Model details
This model is a 17B parameter Mixture-of-Experts architecture designed to balance high-quality performance with operational efficiency. It functions as an image-text-to-text system, utilizing a compressed tensor quantization approach to facilitate effective processing. Its design intent centers on providing a robust tool for creative writing, conversational interactions, and precise visual recognition, including tasks like image reasoning and captioning. By leveraging the transformers library, the model is built to serve as a flexible foundation for developers creating AI assistants and multimodal applications.
The model benefits from conversational fine-tuning, which enhances its utility for interactive applications and complex user-facing tasks. Beyond its primary conversational strengths, the model supports advanced workflows such as synthetic data generation and distillation, allowing its outputs to be used to improve other systems. Its architecture is optimized for both text generation and multimodal understanding, making it a practical choice for developers seeking a balance between performance and resource management. As part of the broader Llama 4 series, it represents a forward-looking approach to integrating visual and textual intelligence into scalable AI solutions.
A provider subscription or plan supersedes token-based pricing for this model.
Llama
The latest AI models from Meta, Llama-4-Scout-17B-16E-Instruct and Llama-4-Maverick-17B-128E-Instruct-FP8, are now available on GitHub Models.
This exact model name is also listed by 4 other providers.