Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
NanoGPT logo

Model details

DeepSeek V4.1 Flash TEE

The model overview is temporarily unavailable.

NanoGPTTEE/deepseek-v4.1-flashdeepseek-flash

Quick Info

Powered by
Provider
NanoGPT
Model key
TEE/deepseek-v4.1-flash
Release date
Sep 10, 2026
Last updated
Sep 10, 2026
Knowledge cutoff
2025-05
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.65
Output token cost
$1.45

Limits

Input tokens
1,048,576 tokens
Output tokens
384,000 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare DeepSeek V4.1 Flash TEE pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about DeepSeek V4.1 Flash TEE

NanoGPT

CoverageBenchmark

DeepSeek released DeepSeek-V4.1-Flash on September 10, 2026, with architectural changes documented in the official model card. The backbone grew from 284B to 552B parameters while active parameters fell from 13B to 8B for input reading and 16B for output generation, reflecting a shift to a Causal Encoder-Decoder archit Benchmark jumps versus V4 Flash include Terminal-Bench 2.1 moving from 82.7 to 90.6, Terminal-Bench 4.0 from 7.0 to 31.2, and DeepSWE v1.1 from 54.4 to 74.2. DeepSeek announced that from 12:00 Beijing time on September 14, 2026 (04:00 UTC), requests to deepseek-v4-pro are routed to V4.1 Flash and billed at Flash rates,

NanoGPT

CoverageAnalysis

DeepSeek V4.1 Flash is described as a low-cost, open-weight, multimodal model for long agent loops, combining a 1M-token context window with up to 384K output tokens and DeepSeek's official API ID deepseek-flash. The model's practical identity is five parts: multimodal image-and-text input, text output, a 552B-paramete Legacy names deepseek-v4-flash and deepseek-v4-flash-vision-exp temporarily route to V4.1 Flash at Flash rates, and from 04:00 UTC on September 14, 2026, DeepSeek routes requests for deepseek-v4-pro to V4.1 Flash until a later V4.1 Pro release. The piece explicitly recommends clients use the exact ID shown by their pro

NanoGPT

CoverageLeaks

DeepSeek V4.1 Flash moved from a two-day beta (carrying the internal ID deepseek-v4.1-flash-expires-on-0910) to a full general release on September 10, 2026, available across DeepSeek's app, web interface, and API under the model name deepseek-flash. DeepSeek published a technical report titled Pushing the Limits of KV The release artifacts include 48 safetensors shards totaling roughly 510 GB that went online the morning of release, letting developers inspect the new architecture directly. The piece also references the prior V4 Flash line and lists competing models served on the same routing platform, providing release-timing contex

Videos about DeepSeek V4.1 Flash TEE

More models around DeepSeek V4.1 Flash TEE