Currently listed through these providers:
Model details
Gemma 4 31B Instruct
Google DeepMind built Gemma 4 31B Instruct as a 30.7 billion parameter dense model that breaks from the trend of ever-larger architectures by delivering strong performance in a compact footprint. The model is natively multimodal, accepting text, images, and video inputs while generating text output, and it includes configurable thinking and reasoning modes that let applications adapt how the model approaches complex problems. Its 262K token context window provides substantial room for analyzing lengthy documents or maintaining extended conversations, and native function calling enables seamless integration into agentic workflows without additional prompting gymnastics.
The Gemma 4 31B family ships under an Apache 2.0 license, making it a practical choice for developers who want to inspect, fine-tune, or deploy the model in private environments without licensing friction. Instruction-tuning shapes the base model into a helpful conversational partner, and the open weights approach has driven significant community adoption with millions of downloads on Hugging Face. The model targets real-world strengths in coding, structured reasoning, and multilingual document understanding across 140 languages, positioning it as a versatile workhorse for applications ranging from autonomous agents to content generation pipelines where transparency and control matter.
Quick Info
Powered by- Provider
- Together AI
- Model key
- google/gemma-4-31B-it
- Release date
- Apr 7, 2026
- Last updated
- Apr 7, 2026
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.39
- Output token cost
- $0.97
Limits
- Output tokens
- 131,072 tokens
- Context window
- 262,144 tokens
Transparent token rates
Compare gemma pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Gemma 4 31B Instruct
No articles yet. Fetch the latest news to show it here.