Tempr
BigGo Finance reports that Qwen3.8-27B, released under Apache 2.0 by Alibaba's Tongyi Qianwen team, passed 1 million Hugging Face downloads within two days and inspired roughly 500 community quantized builds. The 27B dense model uses a hybrid Gated DeltaNet and Gated Attention architecture with a 262K native context window and multi-token prediction (MTP), and runs on consumer GPUs and workstations once quantized. The piece says NVIDIA, AMD, Cerebras, vLLM, and SGLang completed integrations shortly after release, while overseas developers pushed MTP speculative decoding, reasoning-effort tuning, quantization, and Apple Silicon work that yielded over 60% decoding-speed gains on some hardware. It situates Qwen3.8-27B inside broader Tongyi Qianwen claims of surpassing Qwen3.7-Plus and outperforming Claude Opus 4.6 Max on coding, agentic, and multimodal benchmarks, though those results trace to vendor-supplied numbers.