Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
LLM Gateway logo

Model details

Qwen3.8 Max (SCX.ai)

Qwen3.8 Max is described as a flagship 2.4-trillion-parameter mixture-of-experts model in the Qwen3.8 series, built on the architectural foundation of Qwen 3.5 and announced by Qwen as the most capable entry in the family to date. The design combines a massive MoE parameter budget with native visual understanding, allowing the model to read images alongside text without separate vision adapters. Qwen's release notes highlight comprehensive improvements across coding and professional work, framing Qwen3.8 Max as a frontier-tier system aimed at complex software engineering and long-horizon autonomous agent tasks that require sustained reasoning across many steps.

On the SCX.ai route, Qwen3.8 Max is offered with streaming, tool calling, a configurable reasoning budget, JSON-structured output, and an optional web search mode that adds a small per-search surcharge. The provider advertises a one-the cataloged API limit, making the model well suited to very long documents, multi-file codebases, and extended agent traces that need large working memory. With closed weights hosted through SCX.ai and cache pricing that rewards prompt reuse, this route is a practical fit for teams that need top-of-the-line Qwen reasoning and vision capabilities without self-hosting, especially for autonomous coding and research workflows that benefit from long context and persistent caching.

LLM Gatewayscx-ai-gp/qwen3.8-maxqwen

Quick Info

Powered by
Provider
LLM Gateway
Model key
scx-ai-gp/qwen3.8-max
Release date
Aug 3, 2026
Last updated
Aug 3, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$2.00
Output token cost
$6.00

Limits

Output tokens
131,072 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare Qwen3.8 Max (SCX.ai) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.8 Max (SCX.ai)

LLM Gateway

CoverageRelease Notes

DataNorth's September 2, 2026 recap corroborates the QwenCloud release of Qwen3.8-Max-0902, framing it as a reproducible API snapshot of the existing Qwen3.8-Max flagship rather than a new model family. The article restates Alibaba's vendor claims of improved coding for engineering-scale projects, stronger long-horizon The piece adds concrete API-surface details pulled from QwenCloud for the Qwen3.8-Max-0902 snapshot: a 1 million token context window, 128,000 token maximum output, 256,000 token thinking budget, and support for function calling, built-in tools, and structured output. It does not mention the SCX.ai route or LLM Gateway

LLM Gateway

CoverageRelease Notes

DataNorth confirmed that Qwen3.8-Max reached general availability on August 3, 2026, after a mid-July 2026 preview. The model features 2.4 trillion total parameters with 95 billion active per query, a 1-million-token context window, and native multimodal input across text, images, and video. Pricing is set at $2.00 per Alibaba is also releasing Qwen3.8-27B as a second open-weight checkpoint aimed at ordinary on-premise GPU hardware, making Qwen3.8-Max the first model in the Qwen-Max class to have its weights published. This marks Alibaba's return to open-sourcing its top-tier models after keeping several 2026 flagship releases propri

LLM Gateway

CoverageRelease Notes

MarkTechPost reported on August 3, 2026 that Alibaba's Qwen team released Qwen3.8-Max and confirmed that open weights would ship the following week. A second checkpoint, Qwen3.8-27B, is also going open-weights. Qwen3.8-Max is a 2.4-trillion-parameter mixture-of-experts model accepting text, image, and video as input an The release marks Alibaba's continuation of shipping its top-tier models with publicly available weights. Qwen3.8-Max's multimodal input support and massive parameter count position it as Alibaba's most capable model in the Qwen family to date, with deployment options through hosted APIs alongside the upcoming open-wei

LLM Gateway

Coverage

Alibaba Cloud confirmed the August 3, 2026 launch of Qwen3.8-Max, a 2.4-trillion-parameter Sparse MoE flagship with 95 billion active parameters and a 1-million-token context window. The model is now accessible via APIs on Alibaba Cloud Model Studio for global developers, with weights scheduled for release the followin Built upon the foundation of Qwen 3.5, Qwen3.8-Max balances massive scale with inference efficiency through its Sparse MoE and hybrid attention design. The model achieves fourth place in Frontend Code Arena and showed strong autonomous coding capability, completing a real-world software engineering project over a 16-da

LLM Gateway

Coverage

Alibaba officially launched Qwen3.8-Max on August 3, 2026, positioning it as the most powerful model in the Qwen series to date. The model is a multimodal system with 2.4 trillion total parameters using a Sparse Mixture-of-Experts architecture with a hybrid attention mechanism, activating 95 billion parameters per quer Qwen3.8-Max demonstrates frontier-level capabilities in autonomous coding and real-world applications, ranking fourth in Frontend Code Arena. In internal testing, the model autonomously executed a real-world software engineering project over a 16-day period, tasked with creating a self-evolving agent framework from scr

LLM Gateway

CoverageRelease Notes

QwenCloud's official changelog documents a new dated snapshot of the Qwen3.8-Max flagship, published on September 2, 2026 as qwen3.8-max-0902 (alias qwen3.8-max-2026-09-02). The release notes describe upgrades targeting complex engineering-scale coding projects, long-horizon autonomous development, and multi-tool colla Because the subject model is an LLM Gateway-hosted alias routed via SCX.ai for the Qwen3.8-Max family, this first-party Qwen changelog provides the most authoritative evidence about the underlying model behavior a developer would experience through that route. The page does not itself name SCX.ai or LLM Gateway, and it

LLM Gateway

CoverageBenchmark

BenchmarkList documents Qwen3.8 Max's performance across 77 evaluations. The model ranks first of 17 on Tau3-Banking with 51.3% Pass@1, tenth of 34 on GDPval-AA with an Elo score of 1,721, and second of 16 on Workspace-Bench with a 67.7% rubric pass rate. It also achieved an Elo of 1,420 on AA-Briefcase (rank 8 of 56), Benchmark results span dates from August through early September 2026, covering comparisons against Claude Fable 5.1, GPT-5.6 Sol, Kimi K3, and the Qwen3.8-2.4T-A95B open-weight checkpoint. The data provides concrete independent performance evidence for the Qwen3.8 Max family across agentic tool-use, workspace tasks, a

Videos about Qwen3.8 Max (SCX.ai)

More models around Qwen3.8 Max (SCX.ai)