Currently listed through these providers:
Model details
gemma 4 31B turbo TEE
The Gemma 4 31B Turbo TEE represents Google's push to deliver capable open-weight inference within trusted execution environments. As a 31 billion parameter model from the Gemma family, it balances computational efficiency with strong reasoning capabilities. The TEE designation indicates this variant is optimized for secure, isolated execution—a feature that makes it attractive for applications requiring confidentiality guarantees beyond standard cloud deployments. It processes both text and image inputs while generating text outputs, making it versatile for multimodal workflows while maintaining the accessibility that has made the Gemma family popular among developers and researchers.
Available through the Chutes platform, this model integrates into existing AI application pipelines via standard API endpoints. The combination of open weights, substantial context capacity, and tool-calling support positions it well for developers building autonomous agents, research assistants, or custom enterprise applications. The turbo designation suggests Google applied targeted optimizations to improve inference performance relative to base configurations. Its support for structured output and temperature control gives developers fine-grained control over response behavior, enabling everything from precise technical responses to more creative generative tasks within the same model.
Quick Info
Powered by- Provider
- Chutes
- Model key
- google/gemma-4-31B-turbo-TEE
- Release date
- Apr 2, 2026
- Last updated
- Apr 2, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.12
- Output token cost
- $0.37
Limits
- Output tokens
- 65,536 tokens
- Context window
- 131,072 tokens