Currently listed through these providers:
Model details
Gemini 3.5 Flash
Gemini 3.5 Flash is built on the established Gemini 3 Flash reasoning foundation, layering on explicit thinking levels that let developers tune the tradeoff between quality, cost, and latency depending on the task. Google positions it as the model that closes the gap between Flash speed and Pro-level capability, designed specifically to handle agentic workflows and complex coding tasks that previously demanded heavier, more expensive models. The architecture supports text, images, audio, video, and PDFs as inputs, with first-party tooling for function calling, structured output, code execution, and search-as-tool—giving developers a cohesive platform for building autonomous agents and integrated applications.
The model launched at Google I/O 2026 as part of the broader Gemini 3.5 family, slotting into Google's ecosystem alongside enterprise-focused platforms and workspace integrations. Independent benchmarks across nine Appwrite service categories have tested the claim that a mid-tier model can carry workloads previously limited to the Pro tier, with the "high" thinking configuration appearing in most of Google's published performance numbers. For developers prioritizing sustained performance on agentic and coding tasks without the latency overhead of larger models, Gemini 3.5 Flash represents Google's push to make frontier-class reasoning accessible at Flash-scale efficiency.
Quick Info
Powered by- Provider
- OpenRouter
- Model key
- google/gemini-3.5-flash
- Release date
- May 19, 2026
- Last updated
- May 19, 2026
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.50
- Output token cost
- $9.00
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,048,576 tokens
Latest news about Gemini 3.5 Flash
Videos about Gemini 3.5 Flash
Recent tweets and retweets from OpenRouter
More models around Gemini 3.5 Flash
This exact model name is also listed by 22 other providers.