Currently listed through these providers:
Model details
Gemma 4 31B IT
Gemma 4 31B IT is the largest dense variant in the Gemma 4 open-weights family from Google DeepMind, sharing the same research foundation as the Gemini line while remaining available under an Apache 2.0 license. It is purpose-built for advanced reasoning, coding, and agentic workflows, with a dense transformer architecture that distinguishes it from the sparse mixture-of-experts options in the wider family. The model is positioned for deployment on servers and high-end laptops, offering an alternative to larger proprietary models while keeping weights openly accessible.
As a multimodal release, this instruction-tuned variant accepts text and image inputs and produces text output, letting teams prototype vision-language applications alongside language-only tasks. Independent benchmarks cited by SambaNova report 85.2% on MMLU Pro, 89.2% on AIME 2026 without tools, 80.0% on LiveCodeBench v6, 84.3% on GPQA Diamond, and a Codeforces ELO of 2150, signalling competitive reasoning and code performance for its size class. A context window reaching 256K tokens supports long-document analysis, multi-step reasoning chains, and codebases that exceed typical context budgets. Practically, it fits teams that want an open dense model for on-prem or self-hosted reasoning pipelines, agent orchestration, and coding assistants, with tooling and structured output support already covered elsewhere on this page.
Quick Info
Powered by- Provider
- Model key
- gemma-4-31b-it
- Release date
- Apr 2, 2026
- Last updated
- Apr 2, 2026
- Input modalities
- Output modalities
- Capabilities
Limits
- Output tokens
- 32,768 tokens
- Context window
- 262,144 tokens
Latest news about Gemma 4 31B IT
No articles yet. Fetch the latest news to show it here.
Videos about Gemma 4 31B IT
More models around Gemma 4 31B IT
This exact model name is also listed by 27 other providers.