Currently listed through these providers:
Model details
DeepSeek V4 Flash Vision (Exp)
DeepSeek V4 Flash Vision (Exp) extends the DeepSeek Flash family with image understanding while keeping the same lightweight hybrid-attention design that defines its siblings. An experimental build was surfaced on the NVIDIA DGX Spark / GB10 user forum, where community discussion confirmed the model as a live, experimental vision-capable variant of the V4 Flash lineage. Third-party listings describe it as a fast hybrid-attention reasoning model with vision support, indicating the team reused the Flash architecture's efficient attention pattern and added visual processing on top of the text foundation rather than producing a new flagship model.
For practical use, the model is positioned as a low-latency multimodal option for workloads that combine documents, screenshots, or other images with long conversational or analytical context. Independent cataloging sites highlight its the cataloged API limit context window as a defining strength, supporting tasks such as extended document review, code repository analysis with diagrams, and multi-turn agentic flows where both images and large text histories must be retained. Because benchmarks are not yet scored on third-party trackers and no first-party release notes are available, the variant is best understood as an early experimental checkpoint that lets developers prototype vision-enabled applications on the Flash tier's economical, hybrid-attention backbone before more comprehensive evaluations land.
Quick Info
Powered by- Provider
- above.dev
- Model key
- deepseek-v4-flash-vision-exp
- Release date
- Aug 21, 2026
- Last updated
- Aug 21, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.242
- Output token cost
- $0.726
Limits
- Output tokens
- 384,000 tokens
- Context window
- 1,000,000 tokens