Currently listed through these providers:
Model details
Qwen3.7 Flash
Qwen3.7 Flash sits at the Flash tier of Alibaba Cloud's Qwen 3.7 family, positioned as a cost-effective vision-language model that extends strong text capabilities with upgraded multimodal understanding. It accepts text, image, and video inputs and produces text outputs, combining natural-language reasoning with agent-oriented features such as tool calling, computer use, screen reading, GUI operation, and code generation from visual references. The model is described in third-party gateway listings as a "mid-to-high cost-performance Plus" offering within the series, emphasizing real-world scene perception and end-to-end navigation of mobile and desktop applications rather than purely conversational use.
In practical terms, Qwen3.7 Flash is aimed at developers who want multimodal interactive hybrid agent workflows: reading screens and documents, generating code grounded in visual context, and orchestrating multi-step productivity tasks that mix perception with reasoning. Its large context window enables long, mixed-modality sessions such as analyzing extended video transcripts alongside screenshots or handling document-heavy pipelines that need cross-referencing across many pages. Third-party gateway observations show it being routed under identifiers like "alibaba/qwen3.7-flash" on Vercel and a bare "qwen3.7-flash" key on AIHubMix, with tiered pricing that rises for longer inputs, making it most economical for short-to-medium agentic calls while still scaling to the cataloged API limit contexts for heavier retrieval and analysis workloads.
Quick Info
Powered by- Provider
- OrcaRouter
- Model key
- qwen/qwen3.7-flash
- Release date
- Jul 15, 2026
- Last updated
- Jul 15, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.03
- Output token cost
- $0.13
Limits
- Input tokens
- 991,000 tokens
- Output tokens
- 65,536 tokens
- Context window
- 1,000,000 tokens