Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OpenRouter logo

Model details

Gemma 4 31B IT

Gemma 4 31B IT is the largest model in Google DeepMind's Gemma 4 open-weight family, shaped by Gemini 3 research and engineered to push intelligence-per-parameter further than earlier Gemma releases. It is a dense, instruction-tuned, vision-language model, meaning it accepts both text and image inputs, has been post-trained to follow instructions, and processes visual information natively rather than through a bolted-on adapter. This design lineage, drawing directly from frontier Gemini 3 research, signals that the model is intended as a capable general assistant rather than a narrowly specialized tool, while still remaining within the open-weights ecosystem that has defined the Gemma family.

In practical terms, Gemma 4 31B IT is positioned for workloads that demand both reasoning and visual understanding, including coding assistance, agentic workflows that chain tool calls together, structured document extraction, and visual question answering over images or screenshots. Reported benchmark performance is strong for an open model of its size, with approximately 2,150 Codeforces ELO reflecting competitive coding ability, 89.2% on AIME 2026 indicating advanced mathematical reasoning, 84.3% on GPQA Diamond suggesting graduate-level science understanding, and 76.9% on MMMU Pro showing robust multimodal comprehension. These strengths make the model a fitting choice for developers building assistive agents, document understanding pipelines, or reasoning-heavy applications where open deployment and strong baseline performance matter more than absolute frontier scores.

OpenRoutergoogle/gemma-4-31b-itgemma

Quick Info

Powered by
Provider
OpenRouter
Model key
google/gemma-4-31b-it
Release date
Apr 2, 2026
Last updated
Apr 2, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.09
Output token cost
$0.34

Limits

Output tokens
16,384 tokens
Context window
262,144 tokens

Transparent token rates

Compare Gemma 4 31B IT pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemma 4 31B IT

ai&

Coverage

A July 13, 2026 Google Cloud Community article by Vipul Raja provides a deep technical look at Gemma 4 31B IT, explicitly naming the instruction-tuned 31B variant released by Google DeepMind in April 2026 as part of the Gemma 4 family. The piece breaks down the naming convention (Gemma family, 4th generation, 31 billio The same article details key capabilities directly tied to the 31B IT variant: a configurable "thinking" mode for step-by-step problem-solving on complex tasks, native support for function calling and structured output suited to agentic workflows, and an Apache 2.0 license that allows developers and researchers to free

Requesty

CoverageBenchmark

A technical deep dive on Gemma 4 describes the family as multimodal, handling text and image input (with audio on small models) and generating text output, with a context window of up to 256K tokens and multilingual support across 140+ languages. The release includes both pre-trained and instruction-tuned variants. The Google DeepMind's stated positioning is "unprecedented intelligence-per-parameter" purpose-built for advanced reasoning and agentic workflows. Since the first Gemma generation, the ecosystem has seen over 400 million downloads and more than 100,000 community variants. The article frames Gemma 4 against DeepSeek R2, Qwe

Merge Gateway

Coverage

Google Cloud announced Gemma 4 availability on Vertex AI on April 2, 2026, positioning it as the company's "most capable family of open models" built from the same research as Gemini 3 and released under an Apache 2.0 license. The blog explicitly names the Gemma 4 31B dense model as a variant suited for "complex enterp The post also notes that the Gemma 4 26B MoE variant will become fully managed and serverless on Model Garden "over the coming days," and highlights integration with the Agent Development Kit (ADK) for building and deploying AI agents. Enterprise-focused features include deployment across Sovereign Cloud solutions for

OpenRouter

Official sourceBenchmark

OpenRouter hosts Google DeepMind's Gemma 4 31B Instruct as a free API-accessible model, described as a 30.7B dense multimodal model supporting text and image input with text output. The model features a 262K token context window, configurable thinking/reasoning mode, native function calling, and multilingual support ac Benchmark scores reported across providers show GPQA Diamond ranging from 68.3% to 84.3% (Crusoe, Together, DeepInfra, SambaNova, ModelRun all in the 82–84% band), while TAU-Bench scores on the listed providers range from 75.1% to 78.3%. Tool-call error rate sits at ~1.70% and structured-output error rate at ~1.70% on

OpenRouter

Official sourceComparison

OpenRouter provides a direct side-by-side comparison between Gemma 4 26B A4B (MoE) and Gemma 4 31B (Dense), both accessible through the same OpenRouter API by switching model slugs without any new integration. Both variants share an identical 262,144-token context window and are provided by Google. The comparison page Pricing on OpenRouter differs materially: the 26B A4B is listed at $0.042 per million input tokens and $0.22 per million output tokens, while the 31B is priced at $0.09 input / $0.34 output per million tokens. Beyond pricing and the shared context window, the comparison page excerpt does not include benchmark or latenc

OpenRouter

Official sourceComparison

OpenRouter hosts Google's Gemma 4 31B IT (free variant) alongside its broader model catalog, positioning it in comparison buckets including flagship, most affordable, best for code, and reasoning tiers. The free tier competes with OpenRouter-routed alternatives like Claude Fable 5 (batch), GPT-5.5 (batch), Gemini 3.1 P As a free model, Gemma 4 31B provides developers a zero-cost entry point for evaluating Google's 30.7B dense multimodal architecture with its 262K context window. The page also surfaces a promotional banner for GPT 5.6 Terra and Luna at 50% off, though that pertains to other models rather than Gemma 4 itself.

Videos about Gemma 4 31B IT

More models around Gemma 4 31B IT