Sulat.com
AI models
CrossModel logo

Model details

Gemini 3.8 Flash

Gemini 3.8 Flash is positioned by Google as the next iteration in the Gemini 3 model family, building directly on Gemini 3.7 Flash, with stated performance advancements in software engineering and agentic knowledge workflows. The model retains the family pattern of offering customizable effort levels so developers can tune the balance of quality, cost, and latency for their specific use case, which makes it well suited to production scenarios where teams want to dial reasoning depth up or down without switching models. Google explicitly frames the release as delivering next-generation intelligence for agentic workflows, signaling a continued emphasis on tool-driven, multi-step task execution rather than purely conversational interaction.

Alongside the general-purpose release, Google announced a sibling variant called Gemini 3.8 Flash Cyber aimed at cybersecurity applications, which signals the underlying model's flexibility to be specialized for domain-heavy reasoning. The standard Gemini 3.8 Flash model card is published on Google DeepMind's site and documents intended usage, limitations, evaluation results, ethics, and content safety considerations, giving enterprise adopters a transparent reference for responsible deployment. Practically, teams building agentic pipelines that need a fast Flash-class model with controllable reasoning depth, and especially those willing to branch into a cyber-specialized variant, will find this release aligned with iterative, workflow-oriented AI development.

CrossModelgemini/gemini-3.8-flashgemini-flash

Quick Info

Powered by
Provider
CrossModel
Model key
gemini/gemini-3.8-flash
Release date
Sep 2, 2026
Last updated
Sep 2, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.75
Output token cost
$3.75

Limits

Output tokens
65,536 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare Gemini 3.8 Flash pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemini 3.8 Flash

CrossModel

CoverageRelease Notes

Google's official Gemini API release notes confirm that gemini-3.8-flash reached general availability on September 2, 2026. The changelog describes it as Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows, and directs developers t The same official changelog frames the release within a broader Gemini API cadence, including the same-week public preview of Lyria 3.5 for music generation and earlier updates such as agentic video understanding for Gemini 3.7/3.6 Flash and 3.5 Flash-Lite, plus the GA release of gemini-omni-1.1-flash. This positions g

CrossModel

Coverage

Xpert Digital's German-language review frames Gemini 3.8 Flash Cyber as a specialized variant of the Gemini 3.8 series launched September 2, 2026, designed exclusively for detecting security gaps in program code and automatically generating patches. The article names development leads at Google as Tulsee Doshi (Senior The same review cautions that while headline benchmarks and introductory pricing for the freely accessible standard Gemini 3.8 Flash appear competitive, hidden "thinking process" token consumption makes the model approximately 40% more expensive per task than its predecessor at the same per-token rate. This positions G

CrossModel

CoverageBenchmark

BenchLM's aggregator model card for Gemini 3.8 Flash, attributed to Google DeepMind and dated September 2, 2026, lists a 1M token context window, $0.75 input / $3.75 output per million tokens API pricing with cached input at $0.075, and a claimed 327 tok/s throughput with a 10.75s first-token latency. Its composite cap Category percentile breakdowns show Gemini 3.8 Flash at the 97th percentile for both Coding and Knowledge, 93rd for Agentic, and 85th for Multimodal within eligible cohorts. Coding is identified as its strongest eligible category at rank 6, while Agentic ranks 11th, with Agentic noted as its lowest eligible category, p

CrossModel

CoverageBenchmark

DataCamp's analysis reports Gemini 3.8 Flash as Google's most intelligent Flash-tier model released September 2, 2026, and third Flash release in six weeks, priced at $0.75 per 1M input tokens and $3.75 per 1M output tokens through December 31, 2026, then $1.50/$7.50. Benchmarks cited include 90.8% on Terminal-Bench 2. The article also details Gemini 3.8 Flash Cyber, a security-tuned variant restricted to trusted defenders through the new Fairwind Program, which reports a real-world vulnerability discovery rate above 70% and sits on the CWE-Bench Pareto frontier for patching. Both models share the same foundational intelligence, with

CrossModel

CoverageBenchmark

Eesel AI's technical review documents Gemini 3.8 Flash specs: text, image, audio, video and PDF inputs with text output, a 1M token context window, 64K output ceiling, function calling, search-as-tool, and computer use. Distribution surfaces include Google AI Studio, the Gemini API, Android Studio, and Google Antigravi The same review details Gemini 3.8 Flash Cyber as a security variant with deliberately looser mitigations, gated to trusted defenders through the new Fairwind Program, and reproduces key phrasing from Google's model card. Together, the standard 3.8 Flash and Cyber variant form Google's most capable Flash-tier pair to d

CrossModel

CoverageRelease Notes

Ars Technica reports Google released Gemini 3.8 Flash on September 2, 2026, marking the company's third Flash model in just six weeks, framing it as Google's best reasoning and coding Flash-tier model to date. The standard Flash is pitched as a workhorse for agentic tasks and software development, with API access at th The article also covers the parallel launch of Gemini 3.8 Flash Cyber, a variant tuned for vulnerability detection and mitigation on the same foundations, and notes benchmark gains are largest in coding evaluations (topping the DeepSWE leaderboard at lower cost) while OSWorld-2.0 agentic computer use remains behind the

CrossModel

Coverage

Digital Applied's pricing-focused analysis confirms Gemini 3.8 Flash launched September 2, 2026, three weeks after 3.7 Flash, at the identical introductory rate of $0.75 input and $3.75 output per million tokens. The post quotes Google's launch-post footnote directly: introductory pricing expires December 31, 2026, wit The same article cites an independent Artificial Analysis measurement placing Gemini 3.8 Flash at 59 on its Intelligence Index (three points above 3.7 Flash's 56) while finding cost per task rose approximately 40% versus its predecessor at unchanged per-token pricing, indicating higher token consumption per task. This

Videos about Gemini 3.8 Flash

More models around Gemini 3.8 Flash