Currently listed through these providers:
Model details
Llama 4 Maverick 17B Instruct
Llama 4 Maverick 17B Instruct is Meta's contribution to the Llama 4 family, designed as a multimodal language model that can accept both text and image input while producing text output. Its defining architectural choice is a mixture-of-experts design with 128 experts and roughly 17 billion active parameters per forward pass, letting it route tokens through specialized sub-networks rather than activating all weights simultaneously. This MoE structure is what gives the Maverick tier its "high-capacity" label while keeping per-request compute closer to a mid-sized model, making it attractive for teams that want frontier-style flexibility without paying frontier-style inference costs for every prompt.
The model is positioned for teams that want open-weight flexibility with hosted convenience, fitting naturally into self-host experiments, cost-controlled applications, and fine-tune pipelines where the underlying weights can be inspected or adapted. In practice, it is distributed through OpenRouter under the model identifier meta-llama/llama-4-maverick, giving developers a straightforward API path, while the AWS Bedrock model card confirms parallel availability on Bedrock for cloud-native deployments. Independent benchmark indices on Enterprise DNA's directory snapshot list an Artificial Analysis Intelligence Index around 14.3 and a Coding Index around 16.3, alongside competitive Design Arena scores, suggesting Maverick is tuned to be a balanced generalist and coder rather than a single-domain specialist. The combination of open weights, multimodal input, and MoE efficiency makes it a sensible choice for organizations weighing an open-weight model against closed frontier alternatives in total-cost-of-quality comparisons.
Quick Info
Powered by- Provider
- Neon
- Model key
- llama-4-maverick
- Release date
- Apr 5, 2025
- Last updated
- Apr 5, 2025
- Knowledge cutoff
- 2024-08
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.50
- Output token cost
- $1.50
Limits
- Output tokens
- 8,192 tokens
- Context window
- 1,000,000 tokens
Latest news about Llama 4 Maverick 17B Instruct
Videos about Llama 4 Maverick 17B Instruct
More models around Llama 4 Maverick 17B Instruct
This exact model name is also listed by 5 other providers.