Currently listed through these providers:
Model details
DeepSeek V4.1 Flash Marathon
DeepSeek V4.1 Flash Marathon is presented on Subconscious's pricing catalog as a multimodal, high-throughput member of the broader DeepSeek 4.1 family, targeting scenarios that demand both image and text understanding alongside fast response generation. The listing is positioned for agentic and coding-style workloads where the provider highlights a strong cache hit rate on its platform, suggesting the model is tuned to handle repetitive context efficiently. The "Flash" naming convention generally signals a latency-optimized variant within a model family, while the "Marathon" suffix aligns with Subconscious's product branding pattern seen across other hosted models in the same catalog, indicating a deployment-tuned offering rather than a new upstream DeepSeek release.
Because the only supplied evidence is the serving provider's pricing page, deeper technical characterization is limited: no architecture details, parameter counts, training methodology, benchmark results, or context-window specifications from DeepSeek itself are available in the supplied excerpts. What can be said practically is that the model is offered for pay-per-token inference, making it accessible for variable workloads without committed spend. Developers evaluating this model should consult official DeepSeek documentation for authoritative details on capabilities, context handling, and supported use cases beyond the high-level multimodal and high-throughput positioning that Subconscious advertises.
Quick Info
Powered by- Provider
- Subconscious
- Model key
- subconscious/deepseek-v4.1-flash-marathon
- Release date
- Sep 10, 2026
- Last updated
- Sep 10, 2026
- Knowledge cutoff
- 2025-05
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.14
- Output token cost
- $0.28
Limits
- Output tokens
- 384,000 tokens
- Context window
- 5,000,000 tokens