Currently listed through these providers:
Model details
Qwen3.6 27B
A community thread on the NVIDIA Developer Forums announced the release of this 27B-parameter open-weight model from the Qwen family, sparking widespread interest among developers running local inference. The conversation spread across hundreds of posts, reflecting how the release resonated with practitioners experimenting on consumer and prosumer GPU hardware.
Early hands-on reports described smooth multi-hour coding sessions using an FP8 quantized checkpoint served through vLLM, with users praising the default chat template and long-context behavior. The thread also surfaced related variants such as a 35B-A3B mixture-of-experts build, suggesting the family covers a spectrum of deployment sizes for coding, agentic, and general assistant workloads where open weights and flexible tooling matter.
Quick Info
Powered by- Provider
- STACKIT
- Model key
- Qwen/Qwen3.6-27B
- Release date
- Apr 22, 2026
- Last updated
- Apr 22, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.53
- Output token cost
- $0.76
Limits
- Output tokens
- 16,384 tokens
- Context window
- 262,144 tokens