Currently listed through these providers:
Model details
Qwen 3.8 Flash
Qwen 3.8 Flash sits within the broader Qwen family of large language models developed by Alibaba, positioned as a fast, multimodal endpoint aimed at developers building chat, assistant, and tool-driven applications. The cataloged variant is offered through the Vercel AI Gateway, making it accessible alongside other hosted foundation models without requiring self-hosting. Its support for text, image, and PDF inputs combined with text outputs lets teams assemble document-aware assistants and visual question-answering flows in a single API call.
Designed for practical deployment, Qwen 3.8 Flash emphasizes reasoning, tool calling, structured output, and attachment handling, which makes it well suited to agentic workflows where the model must parse documents, invoke external functions, and return machine-readable responses. A context window approaching one million tokens enables long-document analysis and multi-turn conversations with extensive history, while the comfortably large output budget supports detailed generated responses. Together these qualities position the model as a versatile choice for production assistants that need to blend multimodal understanding, extended context, and reliable structured generation in one hosted package.
Quick Info
Powered by- Provider
- Vercel AI Gateway
- Model key
- alibaba/qwen3.8-flash
- Release date
- Aug 26, 2026
- Last updated
- Aug 26, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.16
- Output token cost
- $0.47
Limits
- Output tokens
- 128,000 tokens
- Context window
- 991,000 tokens