Currently listed through these providers:
Model details
Gemini 3.5 Flash
Gemini 3.5 Flash is a mid-tier reasoning model that Google announced at its I/O event on May 19, 2026, framing the release as Pro-level reasoning at Flash-class latency. The model is built on the Gemini the listed price Flash reasoning foundation and introduces explicit thinking levels that let developers tune the trade-off between quality, cost, and response speed, with most of Google's published results drawn from the high thinking configuration. This design targets agentic and coding workloads that previously required the Pro tier, aiming to bring stronger step-by-step reasoning into a lower-latency package suitable for production applications.
According to independent third-party analysis, the model is intended to handle complex multi-step tasks such as code generation, function orchestration, and structured tool use, while remaining responsive enough for interactive use. The reasoning foundation carries over the Flash family's characteristic balance of capability and efficiency, and the tiered thinking controls let teams dial up depth for harder prompts or scale back for simpler queries. For practitioners, this makes Gemini 3.5 Flash a practical fit when you need stronger reasoning than a baseline Flash model but want to keep latency and cost closer to Flash than Pro, especially in agent pipelines, code assistants, and workflows that combine retrieval, tool calls, and multi-turn reasoning.
Quick Info
Powered by- Provider
- Eden AI
- Model key
- vertex/gemini-3.5-flash
- Release date
- May 19, 2026
- Last updated
- May 19, 2026
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.50
- Output token cost
- $9.00
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,048,576 tokens
Latest news about Gemini 3.5 Flash
Videos about Gemini 3.5 Flash
More models around Gemini 3.5 Flash
This exact model name is also listed by 21 other providers.