Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OpenRouter logo

Model details

Gemma 4 31B (free)

Gemma 4 31B Instruct sits inside the Gemma family as a 30.7B dense multimodal release from Google DeepMind, deliberately designed to handle both image and text inputs while producing text outputs. The dense architecture distinguishes it from sibling mixture-of-experts variants in the same lineup, giving it a single, unified parameter set rather than routed experts. Its sizeable 262,144-token context window and 32,768-token output ceiling make it well suited to long documents, extended conversations, and multi-step workflows that benefit from carrying substantial prior context forward. The model also exposes configurable thinking and reasoning modes, allowing applications to dial up deliberation when tasks demand deeper analysis.

Independent benchmark indices place this checkpoint in a pragmatic middle ground rather than at the top of the leaderboard, with a coding index around 43, an agentic index near 14, and an intelligence index close to 29. Those numbers suggest a model best matched to everyday reasoning, instruction-following, and tool-augmented tasks rather than heavily agentic pipelines. Native function calling and multilingual support extend its usefulness for developer workflows that wire it into external tools or serve non-English users. Distributed free of charge through OpenRouter with higher-uptime fallback across providers, it offers a low-friction option for teams exploring the Gemma 4 generation or building prototypes that benefit from open weights.

OpenRoutergoogle/gemma-4-31b-it:freegemma

Quick Info

Powered by
Provider
OpenRouter
Model key
google/gemma-4-31b-it:free
Release date
Apr 2, 2026
Last updated
Apr 2, 2026
Input modalities
Output modalities
Capabilities

Cost

A provider subscription or plan supersedes token-based pricing for this model.

Limits

Output tokens
32,768 tokens
Context window
262,144 tokens

Latest news about Gemma 4 31B (free)

OpenRouter

Official sourceBenchmark

Google's Gemma 4 31B Instruct is available as a free-tier model on OpenRouter under the ID google/gemma-4-31b-it:free. According to the OpenRouter model page, it is a 30.7B-parameter dense multimodal model from Google DeepMind that accepts text and image input and produces text output. It ships with a 256K-token contex The OpenRouter listing shows Google AI Studio as the active provider, with a P50 latency of 1.27s and throughput of 24 tokens per second, alongside 100% uptime and 98.55% inference availability over the trailing three days. The page exposes per-provider benchmark rows for GPQA Diamond (e.g., 84.5% via Together, 84.3% v

OpenRouter

Official sourceBenchmark

Benchmark scores and performance metrics for Google: Gemma 4 31B (free) - Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function calling, and multilingual support

Videos about Gemma 4 31B (free)

More models around Gemma 4 31B (free)