Sulat.com
AI models
Fireworks AI logo

Model details

DeepSeek V4 Flash Vision Exp

DeepSeek V4 Flash Vision Exp extends the DeepSeek V4 Flash line with visual modules and continued training for image understanding. It is presented as a 305-billion-parameter mixture-of-experts base model, combining specialist computation with multimodal input, and is intended for tasks such as describing images, extracting text from screenshots, and interpreting charts. Its text-oriented agent and reasoning performance is described as comparable to the related Flash model, while visual understanding is its main functional extension.

The model is designed for multimodal agents and practical visual analysis, with a 1,040k context window and function calling suited to workflows that combine long text context with images. Reported results include strong agent and chart-oriented benchmark performance, but the evidence supports describing it as an experimental vision-focused variant rather than a universal leader. It is a good fit for developers building image-aware assistants, document and screenshot analysis, chart interpretation, and agents that need to reason over both text and visual inputs.

Fireworks AIaccounts/fireworks/models/deepseek-v4-flash-vision-expdeepseek-flash

Quick Info

Powered by
Provider
Fireworks AI
Model key
accounts/fireworks/models/deepseek-v4-flash-vision-exp
Release date
Aug 21, 2026
Last updated
Aug 21, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.22
Output token cost
$0.66

Limits

Output tokens
384,000 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare DeepSeek V4 Flash Vision Exp pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about DeepSeek V4 Flash Vision Exp

Fireworks AI

Official sourceOfficial

Fireworks AI has listed DeepSeek-V4-Flash-Vision-Exp as Ready on its model library, the first experimental multimodal entry in the DeepSeek-V4 family. According to the Fireworks model page, the model extends DeepSeek-V4-Flash with visual modules and continued training for image understanding, retaining comparable text- The page documents DeepSeek-V4-Flash-Vision-Exp as a 305B-parameter mixture-of-experts base model from Deepseek with 1,040k context, function calling, and image input support, though fine-tuning is not yet available. Serverless pricing is set at $0.22 input / $0.007 cached input / $0.66 output per 1M tokens, with on-de

Videos about DeepSeek V4 Flash Vision Exp

More models around DeepSeek V4 Flash Vision Exp