Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
ZenMux logo

Model details

Step 3.7 Flash (Free)

Step 3.7 Flash is StepFun's response to the demand for fast, agent-friendly multimodal models that can also think step by step. It is built as a sparse Mixture-of-Experts vision-language system with roughly 198 billion total parameters, of which only about 11 billion activate per token, paired with a 1.8 billion parameter vision encoder on top of a 196 billion parameter language backbone. This routing design lets the model deliver reasoning depth on par with much larger dense models while keeping inference cost closer to a small model. Its open weights are published as stepfun-ai/Step-3.7-Flash on Hugging Face, and it accepts text, images, and video alongside standard temperature controls, tool calling, and structured reasoning modes.

In practical use, Step 3.7 Flash is positioned for agentic coding, tool-driven workflows, and multimodal prompts where speed matters as much as accuracy. Independent hands-on testing on an NVIDIA DGX Spark reported a 100% tool-call success rate, a SWE-Bench PRO score of 56.3, and a ClawEval score of 67.1 that placed it ahead of competing flash-tier peers from other labs. With a long the cataloged API limit context window and a free ZenMux routing option, it is a strong fit for developers who want to run vision-aware agents, code assistants, or research pipelines without managing paid inference or running the model themselves, while still benefiting from frontier-style reasoning.

ZenMuxstepfun/step-3.7-flash-free

Quick Info

Powered by
Provider
ZenMux
Model key
stepfun/step-3.7-flash-free
Release date
May 29, 2026
Last updated
May 29, 2026
Knowledge cutoff
2026-03-01
Input modalities
Output modalities
Capabilities

Cost

A provider subscription or plan supersedes token-based pricing for this model.

Limits

Input tokens
256,000 tokens
Output tokens
256,000 tokens
Context window
256,000 tokens

Latest news about Step 3.7 Flash (Free)

ZenMux

Coverage

A blog post on the Kilo platform reports that StepFun released Step 3.7 Flash as its highest-efficiency multimodal Mixture-of-Experts (MoE) model, entering the so-called "Flash wars" alongside Google's Gemini 3.5 Flash class of fast, cost-effective models. The article positions the release as a generational leap over t The same article notes that StepFun released Step 3.7 Flash under a permissive Apache 2.0 license, making the weights openly available for downstream use and self-hosting. It frames the model as part of an industry shift toward sparse-attention and MoE architectures that deliver Flash-tier speed and unit economics for

Videos about Step 3.7 Flash (Free)