Sulat.com
AI models
Umans AI logo

Model details

Umans Flash

Umans Flash is a compact, open-weight coding model built on the Qwen3.6-35B-A3B architecture and quantized to FP8, giving it a deliberately nimble footprint inside Umans AI's lineup. Its design intent is to be the speed-oriented counterpart to heavier siblings like the coder, Kimi, and GLM variants, prioritizing quick turnaround and responsive conversational flow over the deepest long-form reasoning. The model accepts both text and image inputs while producing text output, and it ships with first-class support for attachments, structured output, tool calling, and reasoning, which lets it plug directly into IDE-style agents, command-line coding assistants, and automation pipelines that need vision grounding and code generation in the same turn. It shares the same 262K token context window as its Umans peers, supporting extended prompts and long-horizon tasks while still feeling lightweight in interactive use.

Because Flash sits inside the Qwen family and is delivered with open weights, it inherits the training lineage of that base architecture while Umans layers on its own deployment expertise for efficient serving on its GPU infrastructure. The catalog positions it as a distillation- or throughput-optimized variant of the Qwen backbone rather than a full-scale flagship, and explicitly calls it the fastest option compared to the GLM 5.2, Kimi, and the platform's own larger model. That makes it a practical fit for developers who want frontier-model quality in nimble interactions, and it is exposed both through per-token API access and a flat-rate coding plan aimed at high-volume individual use. For forward-looking workflows, Flash is well suited to agent-driven coding loops, rapid prototyping sessions, and any setting where low latency and inspectable, modifiable weights matter more than maximum single-pass depth.

Umans AIumans-flashqwen

Quick Info

Powered by
Provider
Umans AI
Model key
umans-flash
Release date
Apr 17, 2026
Last updated
Apr 17, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.15
Output token cost
$1.00

Limits

Output tokens
32,768 tokens
Context window
262,144 tokens

Latest news about Umans Flash

No articles yet. Fetch the latest news to show it here.

Videos about Umans Flash

More models around Umans Flash