Vultr
A third-party technical deep dive published on kie.ai on July 14, 2026 documents what independent testers have measured on Qwen 3.6 27B roughly three weeks after the weights landed on Hugging Face. The model is described in the supplied excerpt as a 27-billion-parameter dense transformer with a hybrid linear-attention The deep dive reports concrete developer-relevant numbers: NVFP4 quantization hits MMLU accuracy of 0.8446 and delivers roughly 2.6–2.86× decode speedup over BF16 in a vLLM benchmark. On DGX Spark hardware, community testers measured 28–33 tokens/second single-session throughput on the NVFP4 build via vLLM 0.24.0, whil