Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
LLM Gateway logo

Model details

Qwen3.5 122B A10B (NovitaAI)

The model overview is being prepared.

LLM Gatewaynovita/qwen3.5-122b-a10bqwen

Quick Info

Powered by
Provider
LLM Gateway
Model key
novita/qwen3.5-122b-a10b
Release date
Feb 23, 2026
Last updated
Feb 23, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.40
Output token cost
$3.20

Limits

Output tokens
65,536 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen3.5 122B A10B (NovitaAI) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.5 122B A10B (NovitaAI)

LLM Gateway

Coverage

Alibaba Cloud released Qwen3.5-122B-A10B on February 23, 2026 as a mid-tier multimodal foundation model under Apache 2.0 licensing, featuring 122B total parameters with 10B activated through a 256-expert MoE architecture. The model delivers strong results on MMLU-Pro (86.7%), GPQA Diamond (85.5%), SWE-bench Verified (72.4%), and Terminal-Bench 2.0 (41.6%) per aggregator data. Technical specifications include 262K native context extensible to 1M, grouped-query attention with 32 heads, 2 KV heads at head dimension 256, RoPE positional embedding, and SwiGLU FFN. Self-hosting requires substantial VRAM, such as three RTX 6000 Blackwell GPUs (288GB total) or one MI325X (256GB). Aggregator-listed API pricing stands at $0.40 input and $3.20 output per million tokens, reflecting serving-provider rather than creator pricing.

LLM Gateway

CoverageBenchmark

Qwen3.5-122B-A10B is a native vision-language model built on a hybrid architecture combining linear attention with a sparse mixture-of-experts design for higher inference efficiency. Released February 25, 2026, it features 122B parameters, a 262K context window, and multimodal input support including video, images, and text. It ranks second among Qwen3.5 variants in overall performance per aggregated benchmarks. The model shows an 88% reliability rate across eight benchmarks and excels in hallucination resistance (98.0%, 63rd percentile) and instruction following (75.0%, 80th percentile), with its text capabilities surpassing Qwen3-235B-2507 and visual capabilities exceeding Qwen3-VL-235B. However, it scores 0.0% in general knowledge, coding, reasoning, ethics, and mathematics tests, indicating clear improvement areas. Aggregator-reported provider pricing includes $0.40 input and $3.20 output per million tokens on Novita.

Videos about Qwen3.5 122B A10B (NovitaAI)

More models around Qwen3.5 122B A10B (NovitaAI)