Currently listed through these providers:
Model details
DeepSeek V4 Flash (Alibaba Cloud)
DeepSeek V4 Flash appears in the AIHubMix model catalog as a routed offering under the deepseek-flash family, surfaced through AIHubMix's unified gateway interface. The gateway is documented on the AIHubMix homepage as exposing multiple large language model providers through a single OpenAI-compatible Chat Completions endpoint, which is how developers reach this model alongside other routed options on the platform. Provider documentation for routing configuration and integration sits at docs.aihubmix.com, linked from both the AIHubMix homepage and the model's listing on models.sulat.com.
In practice, the model is positioned as part of a "Flash" tier within the deepseek-flash family on AIHubMix, suggesting a latency- or cost-oriented variant intended for high-throughput routing rather than a flagship tier. Because the listing is delivered through a unified gateway, teams that already standardize on AIHubMix's API surface can adopt the model without managing separate provider credentials or endpoint plumbing. No official architecture details, benchmark numbers, training scale, or context-behavior evidence were available in the supplied sources, so practical evaluation will need to rely on the provider's own documentation and hands-on testing against representative workloads.
Quick Info
Powered by- Provider
- AIHubMix
- Model key
- alicloud-deepseek-v4-flash
- Release date
- Apr 24, 2026
- Last updated
- Apr 24, 2026
- Knowledge cutoff
- 2025-05
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.14
- Output token cost
- $0.28
Limits
- Output tokens
- 384,000 tokens
- Context window
- 1,000,000 tokens
Latest news about DeepSeek V4 Flash (Alibaba Cloud)
Videos about DeepSeek V4 Flash (Alibaba Cloud)
Recent tweets and retweets from AIHubMix
More models around DeepSeek V4 Flash (Alibaba Cloud)
This exact model name is also listed by 19 other providers.