Model details
Gemma-4-31B
Gemma-4-31B is part of a multimodal open-model family developed by Google DeepMind, built with technology traced back to Gemini 3 research and aimed at maximizing intelligence per parameter. The 31B variant sits alongside four other sizes ranging from compact on-device variants up to a larger 26B mixture-of-experts configuration, giving the family a broad deployment footprint from phones and laptops to server environments. Architecturally, the release spans both dense and MoE designs, and Gemma-4-31B is positioned as a capable reasoner with configurable thinking modes, making it well suited to text generation, coding, and structured reasoning workloads that benefit from deliberate step-by-step processing.
The model accepts text and image inputs while producing text outputs, and it runs within a context window of up to roughly 256K tokens, supporting more than 140 languages for multilingual use cases. It is distributed as an open-weights release under the Apache 2.0 license, available in both pre-trained and instruction-tuned variants, with the tuned version accessible through Google AI Studio for quick experimentation. For builders, that combination of a long context window, multimodal grounding, and open weights makes Gemma-4-31B a flexible choice for assistants and pipelines that need to reason over lengthy documents or mixed text-and-image material while remaining customizable for downstream tasks.
Quick Info
Powered by- Provider
- Poe
- Model key
- google/gemma-4-31b
- Release date
- Apr 2, 2026
- Last updated
- Apr 2, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
A provider subscription or plan supersedes token-based pricing for this model.
Limits
- Output tokens
- 8,192 tokens
- Context window
- 262,144 tokens
Latest news about Gemma-4-31B
No articles yet. Fetch the latest news to show it here.