Sulat.com
AI models
OpenRouter logo

Model details

Gemini 3.5 Flash

Gemini 3.5 Flash is built on the established Gemini 3 Flash reasoning foundation, layering on explicit thinking levels that let developers tune the tradeoff between quality, cost, and latency depending on the task. Google positions it as the model that closes the gap between Flash speed and Pro-level capability, designed specifically to handle agentic workflows and complex coding tasks that previously demanded heavier, more expensive models. The architecture supports text, images, audio, video, and PDFs as inputs, with first-party tooling for function calling, structured output, code execution, and search-as-tool—giving developers a cohesive platform for building autonomous agents and integrated applications.

The model launched at Google I/O 2026 as part of the broader Gemini 3.5 family, slotting into Google's ecosystem alongside enterprise-focused platforms and workspace integrations. Independent benchmarks across nine Appwrite service categories have tested the claim that a mid-tier model can carry workloads previously limited to the Pro tier, with the "high" thinking configuration appearing in most of Google's published performance numbers. For developers prioritizing sustained performance on agentic and coding tasks without the latency overhead of larger models, Gemini 3.5 Flash represents Google's push to make frontier-class reasoning accessible at Flash-scale efficiency.

OpenRoutergoogle/gemini-3.5-flashgemini-flash

Quick Info

Powered by
Provider
OpenRouter
Model key
google/gemini-3.5-flash
Release date
May 19, 2026
Last updated
May 19, 2026
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.50
Output token cost
$9.00

Limits

Output tokens
65,536 tokens
Context window
1,048,576 tokens

Latest news about Gemini 3.5 Flash

Videos about Gemini 3.5 Flash

Recent tweets and retweets from OpenRouter

More models around Gemini 3.5 Flash