Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Pioneer logo

Model details

Gemma 4 E4B IT

Gemma 4 E4B IT sits inside Google DeepMind's Gemma 4 family of open-weight models released under the Apache 2.0 license, alongside pre-trained and instruction-tuned variants. The E4B size uses a Mixture-of-Experts design with around 4.5 billion effective parameters, sitting between the smaller E2B and the larger dense and MoE configurations in the lineup. This architecture choice, combined with a hybrid attention approach and Per-Layer Embeddings, is aimed at efficient on-device execution on laptops and phones without sacrificing multimodal capability.

The instruction-tuned variant is built for practical, interactive use: it accepts text, images, and audio natively while producing text, and it adds configurable thinking modes for step-by-step reasoning, native function calling for agentic workflows, and structured conversations through system-prompt support. The model is pretrained on more than 140 languages with out-of-the-box coverage for over 35, and the family targets long-context work, coding assistance, and reasoning-heavy assistants that can run locally. Its blend of open weights, multimodal input, and edge-friendly efficiency makes it a strong fit for developers building private, responsive applications on consumer hardware.

Pioneergoogle/gemma-4-E4B-itgemma

Quick Info

Powered by
Provider
Pioneer
Model key
google/gemma-4-E4B-it
Release date
Apr 2, 2026
Last updated
Apr 2, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.20
Output token cost
$0.20

Limits

Output tokens
32,768 tokens
Context window
32,768 tokens

Transparent token rates

Compare Gemma 4 E4B IT pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemma 4 E4B IT

Pioneer

Coverage

The official Google DeepMind model card on Hugging Face for google/gemma-4-E4B-it describes Gemma 4 as a family of open-weight models released under the Apache 2.0 license, with the E4B variant being one of five distinct sizes (E2B, E4B, 12B, 26B A4B, and 31B). The E4B-it card confirms multimodal text-and-image input, The card situates E4B-it between the phone-sized E2B and the workstation-class 12B, 26B-A4B (MoE), and 31B (dense) variants, targeting laptops and high-end phones as its primary deployment environments. It emphasizes hybrid attention that interleaves local sliding-window attention with full global attention, Per-Layer

Videos about Gemma 4 E4B IT

More models around Gemma 4 E4B IT