Currently listed through these providers:
Model details
Google Gemma 4 31B Instruct
Google Gemma 4 31B Instruct is a compact, 31-billion-parameter dense transformer designed to deliver strong reasoning and task performance without the massive computational demands of frontier-scale models. Rather than relying on sheer parameter count, this model is built around efficient architecture and open access, carrying an Apache 2.0 license that makes it fully open-source and permissive for modification and deployment. It processes long contexts spanning hundreds of thousands of tokens and handles multimodal inputs including text, images, and video, positioning it as a flexible workhorse for complex, multi-turn tasks where smaller, faster models have historically struggled.
In practical testing, this model has demonstrated remarkable capabilities in business reasoning and agentic workflows. Across FoodTruck Bench simulations, it achieved 100% survival and profitability over 30-day runs while maintaining the tightest ROI band on the leaderboard, generating approximately 95,000 reasoning tokens per run internally. Notably, it invoked tools through text-based parsing with zero errors across hundreds of attempts, handling complex tool calling without any native function-calling API. Its cost-per-task is a fraction of leading competitors, often 15 to 40 times cheaper, while consistently allocating capital effectively and following multi-step schemas faithfully. This combination of reliable multi-step reasoning, efficient cost structure, and open design makes it especially well-suited for autonomous agents, complex problem-solving pipelines, and applications where precision and affordability both matter.
Quick Info
Powered by- Provider
- Venice AI
- Model key
- google-gemma-4-31b-it
- Release date
- Apr 3, 2026
- Last updated
- Jun 11, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.12
- Output token cost
- $0.36
Limits
- Output tokens
- 8,192 tokens
- Context window
- 256,000 tokens
Transparent token rates
Compare Google Gemma 4 31B Instruct pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Google Gemma 4 31B Instruct
No articles yet. Fetch the latest news to show it here.