Currently listed through these providers:
Model details
Meta Llama 4 Maverick 17B 128E
This model utilizes a mixture-of-experts architecture, featuring 128 specialized experts to deliver efficient and powerful processing. Designed as a general-purpose, natively multimodal system, it is built to handle complex tasks that require both text and image understanding. By distributing computational load across its expert network, the model achieves high performance in multilingual communication, coding assistance, and visual question answering, making it a versatile tool for developers and researchers working on diverse AI applications.
Engineered to support the evolving needs of agentic systems, this model is optimized for tool-calling and complex reasoning workflows. Its design lineage focuses on providing a balance between high-capacity processing and operational efficiency, allowing it to function effectively in demanding environments. As part of the broader Llama ecosystem, it serves as a robust foundation for building intelligent assistants and automated workflows that require deep integration with external tools and data sources.
Quick Info
Powered by- Provider
- Helicone
- Model key
- llama-4-maverick
- Release date
- Jan 1, 2025
- Last updated
- Jan 1, 2025
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.15
- Output token cost
- $0.60
Limits
- Output tokens
- 8,192 tokens
- Context window
- 131,072 tokens
Transparent token rates
Compare Meta Llama 4 Maverick 17B 128E pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Meta Llama 4 Maverick 17B 128E
No articles yet. Fetch the latest news to show it here.