NanoGPT
DeepSeek released DeepSeek-V4.1-Flash on September 10, 2026, with architectural changes documented in the official model card. The backbone grew from 284B to 552B parameters while active parameters fell from 13B to 8B for input reading and 16B for output generation, reflecting a shift to a Causal Encoder-Decoder archit Benchmark jumps versus V4 Flash include Terminal-Bench 2.1 moving from 82.7 to 90.6, Terminal-Bench 4.0 from 7.0 to 31.2, and DeepSWE v1.1 from 54.4 to 74.2. DeepSeek announced that from 12:00 Beijing time on September 14, 2026 (04:00 UTC), requests to deepseek-v4-pro are routed to V4.1 Flash and billed at Flash rates,