Currently listed through these providers:
Model details
Qwen3.7 Plus
Qwen3.7 Plus extends the Qwen3.7 family into a multimodal interactive hybrid agent that unifies vision and language within a single foundation. According to the official Qwen announcement, it is designed to perceive real-world scenes, read screens and operate GUIs, navigate mobile apps end-to-end, and write code from visual references, while still functioning as a text-driven productivity and coding assistant. This positioning lets it blend GUI and CLI actions inside one agent loop, supporting multi-step workflows from frontend prototyping to complex software engineering rather than limiting it to isolated text tasks.
The model is documented as proprietary, with no published weights for download, and is served through Alibaba Cloud Model Studio. It accepts text, image, and video input and produces text output, with a one-the cataloged API limit that quick-info value window and a maximum output of around sixty-four thousand tokens, giving it substantial room for long agentic sessions and large visual or document attachments. Practical fit centers on agent-driven coding tasks, tool use, GUI automation, and grounded visual question answering, where its hybrid GUI/CLI approach and broad modality coverage offer a more versatile workflow than vision-only or text-only alternatives in the same family.
Quick Info
Powered by- Provider
- EmpirioLabs AI
- Model key
- qwen3-7-plus
- Release date
- Jun 2, 2026
- Last updated
- Jun 12, 2026
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.40
- Output token cost
- $1.60
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,000,000 tokens