Currently listed through these providers:
Model details
Gemma 4 31B MeroMero v2 Thinking
This Gemma 4 family release extends Google's open-weight instruction lineage into a multimodal thinking variant, accepting text and image inputs while returning generated text together with explicit reasoning traces. It is hosted on the NanoGPT platform as an openly available checkpoint, paired with long-context handling and tool-calling support that make it practical for agents, retrieval workflows, and structured document tasks where the model must reason over attached images and lengthy prompts together. The framing positions it as a mid-sized, transparent option for developers who want access to weights alongside runtime reasoning visibility.
Benchmark context for the closely related Gemma 4 31B thinking configuration shows solid graduate-level reasoning at 85.7% on GPQA Diamond, a strong 75.6% on IFBench instruction-following, and 68.3% on AA-LCR long-context reasoning, alongside a measured intelligence index of 29.7 that places it above roughly two-thirds of compared models. Coding performance is moderate, with a coding index of 43.4 and 55% peer ranking, while agentic capability sits lower at a 14.4 agentic index, suggesting a model best suited to expressive dialogue, analytical reasoning, and instruction-heavy multimodal work rather than heavy autonomous coding pipelines.
Quick Info
Powered by- Provider
- NanoGPT
- Model key
- Gemma-4-31B-MeroMero-v2:thinking
- Release date
- Jul 29, 2026
- Last updated
- Aug 24, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.10
- Output token cost
- $0.45
Limits
- Input tokens
- 262,144 tokens
- Output tokens
- 32,768 tokens
- Context window
- 262,144 tokens