Currently listed through these providers:
Model details
Qwen3.7 Max
Qwen3.7 Max is presented by the Qwen Team as "The Agent Frontier," a proprietary, non-open-weight flagship explicitly designed around long-horizon autonomous execution rather than general chat. The release framing emphasizes tasks that span hours and involve hundreds or thousands of tool calls without losing context, making it a fit for coding agents and office productivity automation rather than casual conversation. Its sizeable context window and output ceiling support this positioning, allowing sprawling agentic traces to stay in memory across complex workflows.
The model is exposed through an OpenAI- and Anthropic-compatible API surface, reachable via the Yotta AI Gateway or directly through Alibaba Cloud Model Studio, which lets teams wire it into existing agent stacks without retraining integrations. It also appears on DeepInfra's hosted inference catalog under text generation, broadening deployment options. In practice, Qwen3.7 Max is best suited for production agent pipelines that need sustained reasoning over very long contexts, where its tool-calling depth and large working memory matter more than open-weight flexibility.
Quick Info
Powered by- Provider
- Alibaba Coding Plan (China)
- Model key
- qwen3.7-max
- Release date
- May 21, 2026
- Last updated
- May 21, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $2.50
- Output token cost
- $7.50
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,000,000 tokens