Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Berget.AI logo

Model details

Gemma 4 31B Instruct

Gemma 4 31B Instruct is a dense instruction-tuned model from Google DeepMind that continues the Gemma family's line of open releases under the Apache 2.0 license. It is built around a roughly 30.7 billion parameter architecture and accepts images and text as input while producing text outputs, making it a multimodal generalist rather than a text-only assistant. Within the Gemma 4 generation it sits alongside sibling variants such as Gemma 4 26B A4B, giving teams a choice between a dense 31B model and a mixture-of-experts alternative at a similar scale, which is useful when comparing quality-per-parameter against deployment cost.

Practically, the model is aimed at builders who want transparent weights and flexible behavior for reasoning, tool calling, and structured output workflows. Independent catalog benchmarks place it at a coding index of 43.4, an agentic index of 14.4, and an intelligence index of 29.7, positioning it as a capable mid-size option rather than a frontier leader. On Berget's sovereign infrastructure it runs as an OpenAI-compatible endpoint hosted entirely in Europe, which suits organizations that prioritize data residency, open-source lineage, and integration with agentic coding or document-analysis pipelines without depending on closed providers.

Berget.AIgoogle/gemma-4-31B-itgemma

Quick Info

Powered by
Provider
Berget.AI
Model key
google/gemma-4-31B-it
Release date
Apr 2, 2026
Last updated
Apr 2, 2026
Knowledge cutoff
2025-12
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.275
Output token cost
$0.55

Limits

Output tokens
8,192 tokens
Context window
128,000 tokens

Transparent token rates

Compare Gemma 4 31B Instruct pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemma 4 31B Instruct

Berget.AI

CoverageRelease Notes

NVIDIA's NIM for Vision Language Models 1.7.0 release notes (published August 10, 2026) explicitly list an "updated release of Gemma 4 31B Instruct" alongside the initial release of the sibling Gemma-4-26B-A4B-IT variant, both packaged within the NIM VLM container. This is the strongest model-substantive signal in the Because the announcement comes from NVIDIA's NIM platform rather than from Google DeepMind (the model creator) or Berget.AI specifically, attribution is preserved accordingly: this is a packaging/platform update on the inference-hosting layer, not a new Google DeepMind model release. For developers using the same under

Berget.AI

CoverageBenchmark

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. $0 per million input tokens, $0 per million output tokens. 262,144 token context window, maximum output of 8,192 tokens. Higher uptime with 12 providers. Includes independent benchmarks from Artifici

Videos about Gemma 4 31B Instruct

More models around Gemma 4 31B Instruct