Currently listed through these providers:
Model details
Llama 4 Maverick
Llama 4 Maverick represents Meta's shift toward natively multimodal AI through early fusion, treating text and vision tokens together from the ground up rather than bolting on image understanding afterward. This architectural choice enables more coherent cross-modal reasoning across visual recognition, image reasoning, captioning, and answering questions about visual content. Under the hood, the model uses a Mixture of Experts design with 17 billion activated parameters drawn from a 400-billion-parameter total pool, routing through 128 specialized experts plus one shared expert so each token only activates a fraction of what a comparable dense model would require.
The model collection is designed for enterprise-scale applications where cost efficiency matters, offering high quality at a lower price point than comparable dense alternatives. Its architecture supports practical use cases ranging from building conversational AI assistants that reason about both text and images to creating code generation tools with multilingual support. Benchmarks show strong performance on document understanding and chart reasoning tasks, and an experimental chat version achieved an Elo of 1417 on the LMArena leaderboard. The Llama 4 family also supports leveraging model outputs for synthetic data generation and distillation, opening pathways for organizations to cultivate specialized variants tuned to their specific needs.
Quick Info
Powered by- Provider
- DigitalOcean
- Model key
- llama-4-maverick
- Release date
- Apr 5, 2025
- Last updated
- Apr 30, 2026
- Knowledge cutoff
- 2024-08
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.20
- Output token cost
- $0.696
Limits
- Output tokens
- 16,384 tokens
- Context window
- 128,000 tokens