Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
DigitalOcean logo

Model details

Gemma 4

Gemma 4 31B is Google's open-weight dense model where all 31 billion parameters activate during every inference pass, a design choice that prioritizes output quality over the sparse computational trade-offs of its MoE sibling. Built from Gemini 3 research by Google DeepMind, this instruction-tuned variant targets developers and teams who want a straightforward, high-quality base for coding, reasoning, and agentic workflows without the complexity of selective parameter routing. The hybrid attention architecture interleaves local sliding-window attention with full global attention, unified by consistent Keys and Values across global layers and enhanced by proportional RoPE to handle particularly long context windows efficiently.

The model arrives ready for commercial and non-commercial deployment, supporting function-calling, structured JSON output, and native vision alongside multilingual text generation spanning 140+ languages. These capabilities make it well-suited for building coding assistants, automated agents that handle complex multi-step tasks, and multimodal applications that require flexible output formatting. The dense architecture also suggests strong fine-tuning readiness, giving teams a solid foundation to adapt the model to specific domains without fighting the parameter sparsity of MoE designs.

DigitalOceangemma-4-31B-itgemma

Quick Info

Powered by
Provider
DigitalOcean
Model key
gemma-4-31B-it
Release date
Apr 22, 2026
Last updated
Apr 30, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.18
Output token cost
$0.50

Limits

Output tokens
8,192 tokens
Context window
256,000 tokens

Transparent token rates

Compare Gemma 4 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemma 4

DigitalOcean

Coverage

The Hugging Face paper page hosts the official Gemma 4 Technical Report, authored by the Gemma Team at Google and published July 1, 2026. The abstract describes Gemma 4 as a new generation of open-weight, natively multimodal language models featuring both dense and Mixture-of-Experts architectures ranging from 2.3B to For developers and enterprises evaluating DigitalOcean's hosted Gemma 4 31B-it, the technical report establishes the model's design rationale and capability profile, including its multimodal native support and long-context performance claims. The page does not address DigitalOcean-specific serving, pricing, or API impl

DigitalOcean

CoverageRelease Notes

Google's official Gemma releases page documents the Gemma 4 family rollout that underpins DigitalOcean's hosted Gemma 4 31B-it endpoint. According to the page, Gemma 4 was released on March 31, 2026 in E2B, E4B, 31B and 26B A4B sizes, with the 31B variant explicitly named as part of the lineup. The release notes that G This family-level release log is relevant to the Gemma 4 31B-it served on DigitalOcean's Gradient AI Platform because it confirms the upstream model release, capability set (multimodal inputs, 256K context window), and the explicit inclusion of the 31B size within the Gemma 4 lineup. It does not address DigitalOcean-sp

Videos about Gemma 4

More models around Gemma 4