Sulat.com
AI models
DevPass (LLM Gateway) logo

Model details

Gemini 3.6 Flash

Gemini 3.6 Flash sits in the Gemini family as a speed-oriented, token-efficient tier aimed at developers and end users who need responsive multimodal reasoning. Google DeepMind frames it as the workhorse counterpart in the lineup, emphasizing better coding, knowledge work, and multimodal performance while spending fewer tokens than the prior generation. The release is accompanied by an official model card that documents model information, training and data context, implementation and sustainability footprint, distribution channels, and evaluation results, giving integrators a single reference point for capabilities and known limitations.

Compared with the previous Flash generation, Gemini 3.6 Flash is positioned for measurable efficiency gains, with DeepMind reporting a 17% reduction in output token usage relative to Gemini 3.5 Flash according to the Artificial Analysis Index. That focus on token economy makes it a practical fit for high-volume assistants, automated pipelines, and interactive applications where cost per response and latency matter as much as raw capability. Access is broad: consumers can try it in the Gemini app, while developers can build with it through Google AI Studio, with the same multimodal foundation powering both consumer and programmatic experiences.

DevPass (LLM Gateway)gemini-3.6-flashgemini-flash

Quick Info

Powered by
Provider
DevPass (LLM Gateway)
Model key
gemini-3.6-flash
Release date
Jul 21, 2026
Last updated
Jul 21, 2026
Knowledge cutoff
2026-03
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.75
Output token cost
$3.75

Limits

Output tokens
65,536 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare Gemini 3.6 Flash pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemini 3.6 Flash

DevPass (LLM Gateway)

Coverage

Tech Insider reported on July 23, 2026, that Google shipped three new Gemini models on July 21, with Gemini 3.6 Flash reaching general availability in the Gemini API that same day at $1.50 per million input tokens and $7.50 per million output tokens, per the company's official blog. The article frames Gemini 3.6 Flash The rollout spanned the Gemini API through Google AI Studio and Android Studio, Google Antigravity for developers, the Gemini Enterprise Agent Platform, the Gemini Enterprise app, and the Gemini app itself, with Gemini 3.5 Flash-Lite also rolling into Google Search. The piece characterizes the launch as a distribution

DevPass (LLM Gateway)

CoverageBenchmark

Artificial Analysis published an independent evaluation of the explicitly named variant Gemini 3.6 Flash (high), assigning it an Intelligence Index score of 40 (above the comparable-model median of 29) while consuming a relatively concise 66M output tokens to produce that score. Speed was measured at 216.4 output token The model supports text, image, speech, and video inputs with text output, offers a 1M-token context window (approximately 1500 A4 pages), and has reasoning enabled on this variant, with a non-reasoning sibling noted as possibly available. The page classifies the model as proprietary, released July 2026, and compares i

DevPass (LLM Gateway)

Coverage

Google announced Gemini 3.6 Flash on July 21, 2026, alongside companion models Gemini 3.5 Flash-Lite and Gemini 3.5 Flash Cyber. According to the official Google blog post (by Senior Director of Product Management Tulsee Doshi on behalf of the Gemini team), Gemini 3.6 Flash is positioned as the workhorse model deliveri Gemini 3.6 Flash reached general availability in the Gemini API, Google AI Studio, Android Studio, Antigravity, the Gemini Enterprise Agent Platform, and the Gemini app, while Gemini 3.5 Flash-Lite rolled into Google Search. Google framed the release as targeting token efficiency, lower latency, and reliability for pro

Videos about Gemini 3.6 Flash

More models around Gemini 3.6 Flash