Gemini 3.6 Flash launched on July 21, 2026 as a stable, production-ready model in Google's Gemini API, with the stable model ID gemini-3.6-flash. Google built it on Gemini 3.5 Flash and positioned it as a workhorse model for developers handling code, documents, images, audio, video, and tool-driven workflows, reporting The model accepts text, image, video, audio, and PDF as inputs and outputs text, supporting up to 1,048,576 input tokens and 65,536 output tokens. Google reported that it focuses on coding, knowledge work, visual tasks, and token efficiency, with distribution also covering enterprise and agent development channels. As
Model details
Gemini 3.6 Flash
Gemini 3.6 Flash is a Flash-tier model in the Gemini family, built as the direct successor to Gemini 3.5 Flash and released by Google on July 21, 2026 alongside the more economical 3.5 Flash-Lite and the security-focused 3.5 Flash Cyber pilot. It accepts text, image, video, audio, and PDF inputs while producing text-only output, making it suited for workflows that combine code, documents, charts, and multimedia references. The model inherits the Flash lineage's emphasis on long-context understanding and tool-driven agents, and positions itself as an everyday workhorse rather than a heavyweight reasoning system.
The model raises the bar on several practical fronts compared with its predecessor, including notable gains on SWE-Bench Pro (58.7% versus 55.1%), DeepSWE v1.1 (49% versus 37%), MLE-Bench (63.9% versus 49.7%), and OSWorld-Verified (83.0% versus 78.4%), while roughly doubling long-context retrieval performance on GDM-MRCR v2 at the full 1 million-token depth (54.0% versus 26.6%). Google also reports about 17% fewer output tokens than 3.5 Flash on the Artificial Analysis Index, helping offset cost for tool-heavy or multi-step runs. Gemini 3.6 Flash was superseded by Gemini 3.7 Flash in mid-August 2026 but remains a stable option for teams already standardized on it, especially for coding agents, enterprise document processing, and multimodal pipelines that benefit from its large context window and improved tool-call efficiency.
Quick Info
Powered by- Provider
- Vercel AI Gateway
- Model key
- google/gemini-3.6-flash
- Release date
- Jul 21, 2026
- Last updated
- Jul 21, 2026
- Knowledge cutoff
- 2026-03
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.75
- Output token cost
- $3.75
Limits
- Output tokens
- 64,000 tokens
- Context window
- 1,000,000 tokens
Transparent token rates
Compare Gemini 3.6 Flash pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Gemini 3.6 Flash
Google released Gemini 3.6 Flash on July 21, 2026 as the new default workhorse model in the Gemini family and the successor to Gemini 3.5 Flash, alongside simultaneous launches of Gemini 3.5 Flash-Lite and the access-restricted Gemini 3.5 Flash Cyber. It keeps the 1 million-token context window from its predecessor, ad Gemini 3.6 Flash runs at approximately 280 tokens per second, placing it among the faster models in its class for interactive use. The article notes that Google did not release Gemini 3.5 Pro in this announcement because it fell short of internal expectations on coding and complex reasoning, and that Gemini 4 was tease
Videos about Gemini 3.6 Flash
More models around Gemini 3.6 Flash
This exact model name is also listed by 22 other providers.