Currently listed through these providers:
Model details
Qwen3.6 35B-A3B
Within the Qwen family of language models, Qwen3.6 35B-A3B has drawn early attention from the NVIDIA DGX Spark / GB10 community, where a dedicated forum thread titled "Qwen/Qwen3.6-35B-A3B and FP8 has landed" was opened and quickly attracted sustained discussion among developers fine-tuning local configurations. The thread is tagged under the "agentic-ai" category on the NVIDIA Developer Forums, signaling that practitioners are exploring the model for agent-style workflows and applied reasoning tasks rather than treating it as a general chat baseline. Practitioner interest in an FP8 quantized variant alongside the base release suggests the model is being adopted for memory-efficient local serving on compact accelerator hardware.
For users evaluating the model, the peer-to-peer signal from the DGX Spark community points to a fit for on-device experimentation and tuning, particularly when an FP8 variant is available to ease the memory footprint of the 35B-class weights. The dense thread activity indicates that configuration testing and workflow integration are still active areas of community work, so deployment recipes are best validated against current community findings rather than assumed from older Qwen generations. As an open-weights candidate in the Qwen lineage, it is well suited to developers who want to run and adapt a capable model locally and iterate on agentic pipelines.
Quick Info
Powered by- Provider
- AKI.IO
- Model key
- qwen3.6-35b
- Release date
- Apr 17, 2026
- Last updated
- Apr 17, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.15
- Output token cost
- $0.50
Limits
- Output tokens
- 32,768 tokens
- Context window
- 256,000 tokens
Transparent token rates
Compare Qwen3.6 35B-A3B pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.