Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Nvidia logo

Model details

FLUX.2 Klein 4B

FLUX.2 Klein 4B is a compact image generation and editing model from the broader FLUX.2 family, distilled into a roughly 4 billion parameter rectified flow transformer aimed at responsive, consumer-grade workflows. Rather than a separate editing pipeline, it natively handles text-to-image synthesis, single-reference editing, and multi-reference composition in a single model, removing the need to swap weights when moving between creation and modification tasks. A companion undistilled 4B Base variant is also available for fine-tuning and LoRA training, giving developers room to specialize the model for their own interactive or iterative use cases. Pairing this architecture with the Qwen3 4B text encoder keeps prompt understanding strong relative to the lightweight visual stack, which is unusual at this parameter scale.

The model's defining practical strength is its efficiency footprint: it fits in around 13GB of VRAM, runs on hardware such as the RTX 3090 and RTX 4070, and can reach sub-second generation on modern GPUs, with FP8 and NVFP4 quantized builds further trimming memory for tighter deployments. Released under the permissive Apache 2.0 license, it supports personal, scientific, and commercial use, making it well suited to high-volume or interactive pipelines that need predictable performance without sacrificing quality. Through NVIDIA's NIM for Visual Generative AI, the model is exposed via OpenAI API-compatible endpoints, easing integration into existing stacks. It is a natural fit for teams who want a fast, open-weights visual backbone for rapid prototyping, in-product image features, or batch generation on accessible hardware.

Nvidiablack-forest-labs/flux_2-klein-4bflux

Quick Info

Powered by
Provider
Nvidia
Model key
black-forest-labs/flux_2-klein-4b
Release date
Jan 14, 2026
Last updated
Jan 31, 2026
Knowledge cutoff
2025-06
Input modalities
Output modalities
Capabilities

Cost

A provider subscription or plan supersedes token-based pricing for this model.

Limits

Output tokens
40,960 tokens
Context window
40,960 tokens

Latest news about FLUX.2 Klein 4B

Nvidia

Official sourceRelease Notes

The NVIDIA NIM for Visual Generative AI release notes (v1.4.0) mark the initial addition of black-forest-labs/flux.2-klein-4b as a new supported model on the platform, alongside the release of OpenAI API-compatible endpoints for the FLUX.2-klein-4B NIM. This is the primary first-party changelog entry documenting when t The same notes also track subsequent maintenance, including later security updates for flux.2-klein-4b in v1.5.0 and v1.5.3, confirming that NVIDIA continues to patch the model on NIM after its initial 1.4.0 introduction. The page also summarizes parallel additions and CVE fixes for related models such as flux.1-dev, f

Videos about FLUX.2 Klein 4B

More models around FLUX.2 Klein 4B