Sulat.com
AI models
Eden AI logo

Model details

Gemini 3.5 Flash

Gemini 3.5 Flash is a mid-tier reasoning model that Google announced at its I/O event on May 19, 2026, framing the release as Pro-level reasoning at Flash-class latency. The model is built on the Gemini the listed price Flash reasoning foundation and introduces explicit thinking levels that let developers tune the trade-off between quality, cost, and response speed, with most of Google's published results drawn from the high thinking configuration. This design targets agentic and coding workloads that previously required the Pro tier, aiming to bring stronger step-by-step reasoning into a lower-latency package suitable for production applications.

According to independent third-party analysis, the model is intended to handle complex multi-step tasks such as code generation, function orchestration, and structured tool use, while remaining responsive enough for interactive use. The reasoning foundation carries over the Flash family's characteristic balance of capability and efficiency, and the tiered thinking controls let teams dial up depth for harder prompts or scale back for simpler queries. For practitioners, this makes Gemini 3.5 Flash a practical fit when you need stronger reasoning than a baseline Flash model but want to keep latency and cost closer to Flash than Pro, especially in agent pipelines, code assistants, and workflows that combine retrieval, tool calls, and multi-turn reasoning.

Eden AIvertex/gemini-3.5-flashgemini-flash

Quick Info

Powered by
Provider
Eden AI
Model key
vertex/gemini-3.5-flash
Release date
May 19, 2026
Last updated
May 19, 2026
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.50
Output token cost
$9.00

Limits

Output tokens
65,536 tokens
Context window
1,048,576 tokens

Latest news about Gemini 3.5 Flash

Videos about Gemini 3.5 Flash

More models around Gemini 3.5 Flash