Currently listed through these providers:
Model details
Gemini 3.7 Flash
Gemini 3.7 Flash sits inside the third generation of Google's Gemini family and is positioned as the efficiency tier, sitting between deeper-reasoning Pro variants and lighter Flash-Lite siblings. Official Google documentation frames it as a workhorse aimed at agentic use cases, highlighting strong code generation and terminal-execution behavior that approaches Pro-level quality while keeping per-token cost low. The model's design intent is to deliver multimodal processing across long, multi-step workflows without sacrificing throughput, making it attractive for production systems where latency and token efficiency matter as much as raw reasoning depth.
Beyond raw generation, Gemini 3.7 Flash is engineered for agentic applications, with particular strength in tool use, multi-step reasoning, and structured output for downstream pipelines. Google describes it as bridging the gap between Pro and Flash-Lite tiers, offering multimodal understanding suitable for assistants that handle code, documents, and mixed media in a single context. That balance of capability and efficiency makes it a practical fit for developers building production agents, automated coding assistants, and enterprise workflows that need reliable reasoning at scale rather than the deepest possible single-response analysis.
Quick Info
Powered by- Provider
- Ofox
- Model key
- google/gemini-3.7-flash
- Release date
- Aug 13, 2026
- Last updated
- Aug 13, 2026
- Knowledge cutoff
- 2026-03
- AI SDK package
@ai-sdk/google- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.75
- Output token cost
- $3.75
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,048,576 tokens
Latest news about Gemini 3.7 Flash
Videos about Gemini 3.7 Flash
More models around Gemini 3.7 Flash
This exact model name is also listed by 21 other providers.