Sulat.com
AI models
Requesty logo

Model details

Qwen3.8 2.4T A95B (EU)

Qwen3.8 2.4T A95B is a very large mixture-of-experts language model whose 2.4 trillion total parameters point to a sparse-activation design intended to balance broad world knowledge with efficient compute usage during generation. NVIDIA's technical coverage frames the model around configurable reasoning, signaling that the same weights can switch between lighter, low-latency responses and deeper step-by-step inference depending on the caller's needs. The open-weight availability means the same checkpoint can be self-hosted on suitable accelerator infrastructure or routed through managed services, and the Qwen family background suggests a lineage of multilingual pre-training with emphasis on instruction following, code, and analytical text.

The model's intended use spans long-context research assistance, multi-step agentic workflows, and structured analysis tasks where configurable reasoning is valuable. Independent benchmark coverage positions Qwen3.8 2.4T A95B against leading proprietary frontier systems on comparative evaluations, suggesting the model is competitive on reasoning-heavy suites while remaining usable for everyday chat and tool-augmented work. Practically, the very large active expert count makes this checkpoint a strong fit for organizations that need top-tier analytical quality and are willing to pay for the inference cost, while smaller workloads may be better served by lighter Qwen variants.

Requestyqwen3.8-2.4T-A95B@euqwen

Quick Info

Powered by
Provider
Requesty
Model key
qwen3.8-2.4T-A95B@eu
Release date
Aug 12, 2026
Last updated
Aug 12, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$2.50
Output token cost
$6.00

Limits

Output tokens
262,144 tokens
Context window
1,000,000 tokens

Latest news about Qwen3.8 2.4T A95B (EU)

No articles yet. Fetch the latest news to show it here.

Videos about Qwen3.8 2.4T A95B (EU)

More models around Qwen3.8 2.4T A95B (EU)