Currently listed through these providers:
Model details
DeepSeek V4 Flash 0731
DeepSeek V4 Flash 0731 is a community-discussed open-weight language model that surfaced in NVIDIA DGX Spark forums on the same day it was dated, with threads referencing the "v4 flash 0731" naming convention alongside an "ai index = 50" metric that placed it just below GLM 5.2's reported score of 51. Because the weights were announced as openly available and a parallel thread introduced a GGUF-format build under the DGX Spark Projects category, the model appears designed for local experimentation on consumer and workstation-class hardware rather than purely cloud-hosted inference. The "Flash" suffix in the family name, combined with the lightweight GGUF packaging, points toward a derivative intended for faster iteration, lower-memory deployment, and agent-style use cases, as reflected by the "agentic-ai" tag attached to the GGUF announcement thread.
In practical terms, this release positions DeepSeek V4 Flash 0731 as a lightweight entry point within the broader DeepSeek lineup, suited to developers who want to run an open model locally on DGX Spark-class machines and compare it against other recent open releases such as GLM 5.2. The availability of community-shared GGUF weights on the announcement date lowers the barrier to local evaluation, while the "new model" framing in the forum suggests it represents a refreshed checkpoint rather than a wholesale architectural redesign. For users evaluating agent-style workflows or constrained-hardware deployments, the combination of open distribution, a compact Flash-tier footprint, and immediate availability in a quantization-friendly format makes this checkpoint a reasonable candidate for hands-on testing against similarly scaled open models.
Quick Info
Powered by- Provider
- Perplexity Agent
- Model key
- deepseek/deepseek-v4-flash-0731
- Release date
- Jul 31, 2026
- Last updated
- Jul 31, 2026
- Knowledge cutoff
- 2025-05
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.13
- Output token cost
- $0.26
Limits
- Output tokens
- 384,000 tokens
- Context window
- 1,000,000 tokens
Latest news about DeepSeek V4 Flash 0731
Videos about DeepSeek V4 Flash 0731
Recent tweets and retweets from Perplexity Agent
More models around DeepSeek V4 Flash 0731
This exact model name is also listed by 38 other providers.