Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Crusoe logo

Model details

Gemma 4 31B IT

Gemma 4 31B IT is the largest model in Google DeepMind's Gemma 4 open-weight family, shaped by Gemini 3 research and engineered to push intelligence-per-parameter further than earlier Gemma releases. It is a dense, instruction-tuned, vision-language model, meaning it accepts both text and image inputs, has been post-trained to follow instructions, and processes visual information natively rather than through a bolted-on adapter. This design lineage, drawing directly from frontier Gemini 3 research, signals that the model is intended as a capable general assistant rather than a narrowly specialized tool, while still remaining within the open-weights ecosystem that has defined the Gemma family.

In practical terms, Gemma 4 31B IT is positioned for workloads that demand both reasoning and visual understanding, including coding assistance, agentic workflows that chain tool calls together, structured document extraction, and visual question answering over images or screenshots. Reported benchmark performance is strong for an open model of its size, with approximately 2,150 Codeforces ELO reflecting competitive coding ability, 89.2% on AIME 2026 indicating advanced mathematical reasoning, 84.3% on GPQA Diamond suggesting graduate-level science understanding, and 76.9% on MMMU Pro showing robust multimodal comprehension. These strengths make the model a fitting choice for developers building assistive agents, document understanding pipelines, or reasoning-heavy applications where open deployment and strong baseline performance matter more than absolute frontier scores.

Crusoegoogle/gemma-4-31b-itgemma

Quick Info

Powered by
Provider
Crusoe
Model key
google/gemma-4-31b-it
Release date
Apr 2, 2026
Last updated
Apr 2, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.14
Output token cost
$0.40

Limits

Output tokens
32,768 tokens
Context window
262,144 tokens

Transparent token rates

Compare Gemma 4 31B IT pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemma 4 31B IT

ai&

Coverage

A July 13, 2026 Google Cloud Community article by Vipul Raja provides a deep technical look at Gemma 4 31B IT, explicitly naming the instruction-tuned 31B variant released by Google DeepMind in April 2026 as part of the Gemma 4 family. The piece breaks down the naming convention (Gemma family, 4th generation, 31 billio The same article details key capabilities directly tied to the 31B IT variant: a configurable "thinking" mode for step-by-step problem-solving on complex tasks, native support for function calling and structured output suited to agentic workflows, and an Apache 2.0 license that allows developers and researchers to free

Requesty

CoverageBenchmark

A technical deep dive on Gemma 4 describes the family as multimodal, handling text and image input (with audio on small models) and generating text output, with a context window of up to 256K tokens and multilingual support across 140+ languages. The release includes both pre-trained and instruction-tuned variants. The Google DeepMind's stated positioning is "unprecedented intelligence-per-parameter" purpose-built for advanced reasoning and agentic workflows. Since the first Gemma generation, the ecosystem has seen over 400 million downloads and more than 100,000 community variants. The article frames Gemma 4 against DeepSeek R2, Qwe

Merge Gateway

Coverage

Google Cloud announced Gemma 4 availability on Vertex AI on April 2, 2026, positioning it as the company's "most capable family of open models" built from the same research as Gemini 3 and released under an Apache 2.0 license. The blog explicitly names the Gemma 4 31B dense model as a variant suited for "complex enterp The post also notes that the Gemma 4 26B MoE variant will become fully managed and serverless on Model Garden "over the coming days," and highlights integration with the Agent Development Kit (ADK) for building and deploying AI agents. Enterprise-focused features include deployment across Sovereign Cloud solutions for

Crusoe

CoverageBenchmark

OpenRouter identifies Gemma 4 31B Instruct as a 30.7B dense multimodal model with text and image input, text output, a 256K-token context window, configurable thinking, native function calling, and multilingual support across more than 140 languages. Its free variant was released on April 2, 2026. The provider table reports a Crusoe-hosted score of 84.3% on GPQA Diamond and 78.3% on TAU-Bench. OpenRouter also shows 100% three-day uptime and 98.53% availability over August 31 through September 3, while noting that requests can be rerouted to another healthy provider when upstream failures occur.

Videos about Gemma 4 31B IT

More models around Gemma 4 31B IT