Sulat.com
AI models
OpenRouter logo

Model details

Gemini 3.6 Flash

Gemini 3.6 Flash is a Flash-tier model in the Gemini family, built as the direct successor to Gemini 3.5 Flash and released by Google on July 21, 2026 alongside the more economical 3.5 Flash-Lite and the security-focused 3.5 Flash Cyber pilot. It accepts text, image, video, audio, and PDF inputs while producing text-only output, making it suited for workflows that combine code, documents, charts, and multimedia references. The model inherits the Flash lineage's emphasis on long-context understanding and tool-driven agents, and positions itself as an everyday workhorse rather than a heavyweight reasoning system.

The model raises the bar on several practical fronts compared with its predecessor, including notable gains on SWE-Bench Pro (58.7% versus 55.1%), DeepSWE v1.1 (49% versus 37%), MLE-Bench (63.9% versus 49.7%), and OSWorld-Verified (83.0% versus 78.4%), while roughly doubling long-context retrieval performance on GDM-MRCR v2 at the full 1 million-token depth (54.0% versus 26.6%). Google also reports about 17% fewer output tokens than 3.5 Flash on the Artificial Analysis Index, helping offset cost for tool-heavy or multi-step runs. Gemini 3.6 Flash was superseded by Gemini 3.7 Flash in mid-August 2026 but remains a stable option for teams already standardized on it, especially for coding agents, enterprise document processing, and multimodal pipelines that benefit from its large context window and improved tool-call efficiency.

OpenRoutergoogle/gemini-3.6-flashgemini-flash

Quick Info

Powered by
Provider
OpenRouter
Model key
google/gemini-3.6-flash
Release date
Jul 21, 2026
Last updated
Jul 21, 2026
Knowledge cutoff
2026-03
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.75
Output token cost
$3.75

Limits

Output tokens
65,536 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare Gemini 3.6 Flash pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemini 3.6 Flash

Google

CoverageBenchmark

Gemini 3.6 Flash launched on July 21, 2026 as a stable, production-ready model in Google's Gemini API, with the stable model ID gemini-3.6-flash. Google built it on Gemini 3.5 Flash and positioned it as a workhorse model for developers handling code, documents, images, audio, video, and tool-driven workflows, reporting The model accepts text, image, video, audio, and PDF as inputs and outputs text, supporting up to 1,048,576 input tokens and 65,536 output tokens. Google reported that it focuses on coding, knowledge work, visual tasks, and token efficiency, with distribution also covering enterprise and agent development channels. As

Google

CoverageRelease Notes

Google released Gemini 3.6 Flash on July 21, 2026 as the new default workhorse model in the Gemini family and the successor to Gemini 3.5 Flash, alongside simultaneous launches of Gemini 3.5 Flash-Lite and the access-restricted Gemini 3.5 Flash Cyber. It keeps the 1 million-token context window from its predecessor, ad Gemini 3.6 Flash runs at approximately 280 tokens per second, placing it among the faster models in its class for interactive use. The article notes that Google did not release Gemini 3.5 Pro in this announcement because it fell short of internal expectations on coding and complex reasoning, and that Gemini 4 was tease

Videos about Gemini 3.6 Flash

More models around Gemini 3.6 Flash