Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Regolo AI logo

Model details

Gemma 4 31B IT

Gemma 4 31B IT is the largest variant in Google DeepMind's Gemma 4 family of open models, a lineage explicitly framed as being built from Gemini 3 research and technology with the goal of maximizing intelligence per parameter. Independent coverage describes it as a dense, instruction-tuned, vision-language model, combining text and image understanding with post-training aimed at following user intent. The 31B dense design positions it above the smaller E2B and E4B variants in the family, which are aimed at mobile and edge use cases, making the 31B IT version the family's flagship for higher-capacity deployments.

In practical terms, the model is aimed at developers who need strong reasoning, coding assistance, document understanding, and agent-style workflows in a single open-weight model. Reporting from a third-party inference provider highlights benchmark results that include a 2,150 Codeforces ELO for competitive programming, 89.2% on AIME 2026 for mathematical reasoning, 84.3% on GPQA Diamond for graduate-level science questions, and 76.9% on MMMU Pro for multimodal understanding. That same provider reports top rankings from Artificial Analysis for output speed, time-to-first-token, and end-to-end response times on its serving of the model, suggesting a practical fit for latency-sensitive applications such as coding copilots, document extraction pipelines, and conversational agents that mix text and images.

Regolo AIgemma4-31bgemma

Quick Info

Powered by
Provider
Regolo AI
Model key
gemma4-31b
Release date
Apr 2, 2026
Last updated
Apr 2, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.46
Output token cost
$2.42

Limits

Output tokens
100,000 tokens
Context window
100,000 tokens

Transparent token rates

Compare Gemma 4 31B IT pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemma 4 31B IT

Regolo AI

CoverageBenchmark

OpenRouter's listing confirms Gemma 4 31B IT as Google DeepMind's 30.7B-parameter dense multimodal model with text and image input and text output, an Apache 2.0 license, and a 256K (listed 262K) token context window. Features include configurable thinking/reasoning mode, native function calling, multilingual support a Per-provider benchmark rows show GPQA Diamond scores ranging from 68.3% (DeepInfra Ultra) up to 84.5% (Together), with Tau-Bench scores in the 75–79% band across DeepInfra, SambaNova, Crusoe, and ModelRun. Tool-call error rate on Google AI Studio averages 2.20%, and the cache hit rate averages 10.44%. OpenRouter also s

Regolo AI

CoverageBenchmark

AlphaSense published results from a 245-task financial information analysis benchmark evaluating GPT-5.6 Sol, Claude Haiku 4.5, Claude Sonnet 5, Claude Opus 4.8, Claude Opus 5, Kimi K3, GLM-5.2, Inkling, and Gemma 4 31B. The open-weight Gemma 4 31B matched Claude Sonnet 5 in accuracy while costing roughly 1/40th as muc The benchmark compared two serving strategies: feeding all relevant information directly into the model versus using AlphaSense Search to extract targeted context first. The retrieval-augmented approach reduced per-task cost substantially, which is the main reason Gemma 4 31B's per-task median cost landed so far below

Regolo AI

CoverageBenchmark

Google DeepMind released the Gemma 4 family on April 2, 2026, and Modular became a day-zero launch partner hosting all variants on Modular Cloud powered by MAX. Gemma 4 31B is a 31-billion-parameter dense multimodal model with a redesigned architecture and a 256K context window, natively supporting text, images, and vi On NVIDIA B200, Modular reports that Gemma 4 running on MAX achieves roughly 15% higher throughput than vLLM (vllm-0.18.2rc1.dev7) with no accuracy degradation. Reported MAX (B200) scores include MMLU Pro 84.72%, GSM8K Llama COT 95.94%, and ChartQA 84.69%, with comparable MAX (MI355) figures for AMD. The same MAX-power

Videos about Gemma 4 31B IT

More models around Gemma 4 31B IT