Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OpenRouter logo

Model details

Step 3.5 Flash

Step 3.5 Flash is a foundation model built on a sparse Mixture of Experts architecture, activating only 11B of its 196B total parameters per token. This design targets high inference efficiency, with reported generation speeds of roughly 100–350 tokens per second. StepFun describes it as engineered for frontier reasoning and agentic capabilities while keeping the active compute footprint small, a pattern that places it among the wave of sparse Chinese MoE models emphasizing speed at low active parameter counts.

The model is released under an open weights license and is accompanied by a paper, Hugging Face and ModelScope repositories, GitHub code, and a public chat space, making it practical for builders who want to inspect or deploy it locally. StepFun highlights the model's "intelligence density," claiming its reasoning depth rivals top-tier proprietary systems despite the limited active parameter count. A subsequent open-source drop of the Base checkpoint plus Midtrain artifacts and the SteptronOSS training stack extends the line further, giving downstream teams continuation-pretraining flexibility that is unusual for a release of this scale.

OpenRouterstepfun/step-3.5-flash

Quick Info

Powered by
Provider
OpenRouter
Model key
stepfun/step-3.5-flash
Release date
Jan 29, 2026
Last updated
Feb 13, 2026
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.10
Output token cost
$0.30

Limits

Output tokens
65,536 tokens
Context window
262,144 tokens

Latest news about Step 3.5 Flash

OpenRouter

Official sourceBenchmark

Step 3.5 Flash is StepFun's most capable open-source foundation model. $0.10 per million input tokens, $0.30 per million output tokens. 262,144 token context window, maximum output of 65,536 tokens. Higher uptime with 4 providers. Includes independent benchmarks from Artificial Analysis.

StepFun (China)

Coverage

StepFun released Step 3.5 Flash on February 5, 2026, as a sparse Mixture-of-Experts model with 196B total parameters and only 11B active parameters, claiming frontier-level reasoning capability while generating at 100–350 tokens per second, according to a ThursdAI release index that links to the StepFun X announcement StepFun followed up on March 5, 2026 by open-sourcing Step 3.5 Flash Base and Midtrain checkpoints, an unusually open release that includes the SteptronOSS training stack on GitHub alongside the weights, giving builders continuation-pretraining flexibility under an Apache-2 oriented license. The same index links to the

OpenRouter

Official sourceBenchmark

Benchmark scores and performance metrics for StepFun: Step 3.5 Flash - Step 3.5 Flash is StepFun's most capable open-source foundation model. Built on a sparse Mixture of Experts (MoE) architecture, it selectively activates only 11B of its 196B parameters per token. It is a reasoning model that is incredibly speed effi

OpenRouter

Official sourceBenchmark

Step 3.5 Flash is StepFun's most capable open-source foundation model. 256,000 token context window. Higher uptime with 3 providers. Includes independent benchmarks from Artificial Analysis.

Videos about Step 3.5 Flash