Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OpenRouter logo

Model details

Gemma 4 26B A4B (free)

Gemma 4 26B A4B is a Mixture-of-Experts model from Google DeepMind that takes an unconventional path to efficiency. Rather than scaling up a dense model, it distributes computation across specialized experts, activating only 3.8 billion of its 25.2 billion total parameters per forward pass. This architectural choice lets the model punch above its weight class, delivering quality that rivals dense models roughly twice its size while keeping inference costs dramatically lower. The instruction-tuned setup extends beyond text, accepting images and video alongside prompts, with native function calling that enables it to plug directly into external tools and APIs. Its 256K context window accommodates long documents and extended conversations, while a configurable thinking mode lets users dial in how much deliberate reasoning they want from the model.

Built on the same research foundation as Gemini 3, Gemma 4 draws from Google's frontier model work despite its smaller footprint. The Apache 2.0 license removes barriers for commercial and research applications alike, and the open-weight availability means developers can inspect, fine-tune, or self-host the model without vendor constraints. In practice, Gemma 4 A4B is showing up in AI agent frameworks that need reliable tool use and structured output, as well as in developer workflows that demand fast, affordable inference without sacrificing reasoning depth. The free tier removes another barrier to experimentation, making it accessible for rapid prototyping and smaller deployments where budget constraints previously limited access to capable language models.

OpenRoutergoogle/gemma-4-26b-a4b-it:freegemma

Quick Info

Powered by
Provider
OpenRouter
Model key
google/gemma-4-26b-a4b-it:free
Release date
Apr 2, 2026
Last updated
Apr 2, 2026
Input modalities
Output modalities
Capabilities

Cost

A provider subscription or plan supersedes token-based pricing for this model.

Limits

Output tokens
32,768 tokens
Context window
262,144 tokens

Latest news about Gemma 4 26B A4B (free)

OpenRouter

Official sourceComparison

The OpenRouter comparison landing page explicitly lists the subject variant `Gemma 4 26B A4B (free)` as a Google-provided model with a 262,144-token context window and a free price, confirming the model's continued availability through the OpenRouter API and its positioning alongside flagship, coding, low-cost, and ima Beyond the catalog summary, the page surfaces only navigation UI to other model comparisons (Claude Fable 5.1, Gemini 3.1 Pro Preview Custom Tools, GPT-6 Astra, FLUX.2 Pro, Nano Banana 2, GPT Image 2.5 Sunburst, etc.) and does not add new benchmark numbers, release notes, or feature changes specific to the Gemma 4 26B

OpenRouter

Official sourceBenchmark

The OpenRouter model page explicitly names the subject variant `google/gemma-4-26b-a4b-it:free` and describes Gemma 4 26B A4B IT as an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind with 25.2B total parameters and 3.8B active per token, aiming for near-31B quality at lower compute. It supports mu The page also reports provider-level routing telemetry and benchmark results for the free variant, including a P50 latency of 0.89s and throughput of 36 tok/s best across providers, GPQA Diamond scores by provider (SiliconFlow 76.7%, Parasail/NextBit/Cloudflare 75.4%, auto-routing 70.0%) and a TAU-Bench of 69.8% (NextB

Videos about Gemma 4 26B A4B (free)

More models around Gemma 4 26B A4B (free)