Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
StepFun Step Plan (China) logo

Model details

Step 3.5 Flash

Step 3.5 Flash is StepFun's open-source foundation model built on a sparse Mixture-of-Experts architecture, which selectively activates only 11 billion of its 196 billion total parameters per token. This approach prioritizes what StepFun describes as "intelligence density" over brute parameter scale, allowing the model to deliver reasoning depth while keeping per-token computational cost manageable. Published weights are hosted on Hugging Face under the stepfun-ai organization, giving researchers and developers direct access to the model. The design emphasis on activating a small expert subset per token is central to its performance profile and distinguishes it from dense transformer counterparts.

Positioned as a reasoning model optimized for agent and code workflows, Step 3.5 Flash targets scenarios that demand both depth of inference and sustained throughput at long context lengths. StepFun's Step Plan platform exposes the model via dedicated API paths, and documentation describes it as high-speed inference tuned for intelligent agent applications. A derivative checkpoint, step-3.5-flash-2603, builds on this foundation with further optimizations for high-frequency agent scenarios, offering improved token efficiency, faster inference, and an optional low-reasoning mode for cost-sensitive deployments. This variant family suggests a roadmap focused on practical agent deployment rather than purely academic benchmark leadership.

StepFun Step Plan (China)step-3.5-flash

Quick Info

Powered by
Provider
StepFun Step Plan (China)
Model key
step-3.5-flash
Release date
Jan 29, 2026
Last updated
Feb 13, 2026
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities

Limits

Input tokens
256,000 tokens
Output tokens
256,000 tokens
Context window
256,000 tokens

Latest news about Step 3.5 Flash

StepFun Step Plan (China)

Coverage

ThursdAI's StepFun releases index explicitly names Step 3.5 Flash as a new open-weight model released on February 5, 2026, describing it as a 196B sparse MoE with only 11B active parameters that claims frontier-level reasoning while generating at 100–350 tokens per second. The index also points to primary-source links The same index documents a related March 5, 2026 open-source release of Step 3.5 Flash Base and Midtrain checkpoints, bundled with training artifacts and the SteptronOSS training stack on GitHub — described as an unusually open release that includes the underlying training pipeline alongside the weights. These entries

Videos about Step 3.5 Flash