Currently listed through these providers:
Model details
Qwen3 235B A22B Thinking 2507 TEE
As a member of the Qwen3 family, the 235B A22B Thinking 2507 TEE variant is positioned for extended chain-of-thought workloads, where the A22B designation indicates an active-parameter design intended to deliver deep reasoning quality while keeping inference efficient relative to a fully dense model of the same total size. The "Thinking" labeling in the model name signals that it is tuned to spend additional internal deliberation on multi-step problems before producing a final answer, making it a natural fit for analytical tasks such as code synthesis, mathematical proofs, document-grounded research, and planning workflows. Hosting on Chutes exposes the model through an API that supports tool calling and structured output, so the reasoning capability can be chained into agentic pipelines where intermediate steps need to invoke external functions or emit machine-readable responses.
A practical strength of this deployment is the very large 262,144-token context window paired with an equally large maximum output, which lets developers pass entire codebases, lengthy transcripts, or full document collections in a single prompt and still receive substantive long-form responses. Compared with the smaller Qwen3.6 27B TEE and the larger GLM 5.2 TEE offered alongside it on Chutes, the 235B A22B Thinking variant sits in a middle ground aimed at users who want flagship-class reasoning depth without paying the highest per-token rates in the catalog. Open-weight availability adds flexibility for teams that may want to self-host the base model later while using the Chutes endpoint for prototyping, and temperature control plus structured output support make it straightforward to tune response style and integrate with downstream parsers.
Quick Info
Powered by- Provider
- Chutes
- Model key
- Qwen/Qwen3-235B-A22B-Thinking-2507-TEE
- Release date
- Jul 1, 2025
- Last updated
- Jun 21, 2026
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.2989
- Output token cost
- $1.1957
Limits
- Output tokens
- 262,144 tokens
- Context window
- 262,144 tokens