Sulat.com
AI models
Kilo Gateway logo

Model details

Gemini 3.8 Flash

Gemini 3.8 Flash is positioned as the next step in Google's Gemini 3 lineup, evolving from Gemini 3.7 Flash with a focus on stronger software engineering and agentic knowledge workflows. Google DeepMind frames the Flash tier as a balance point, retaining customizable effort levels so developers can tune the trade-off between answer quality, latency, and cost per call. That emphasis on controllable effort makes the model a natural fit for production pipelines that need predictable response times while still benefiting from reasoning-capable behavior on more demanding prompts.

Official documentation confirms both an HTML model card and a PDF model card hosted by Google DeepMind, alongside a dedicated developer's guide within the Gemini Enterprise Agent Platform documentation tree, all dated to the same September 2026 release window. These artifacts are intended to communicate essential information about the model, including known limitations, mitigations, safety performance, and evaluations, with the publisher noting that the cards can be refreshed as the model is improved. For practitioners, that combination of official model card, PDF reference, and Cloud-side developer guide signals a release that is meant to be integrated into real applications, especially agent-style systems that combine reasoning, tool use, and structured outputs.

Kilo Gatewaygoogle/gemini-3.8-flashgemini-flash

Quick Info

Powered by
Provider
Kilo Gateway
Model key
google/gemini-3.8-flash
Release date
Sep 2, 2026
Last updated
Sep 2, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.75
Output token cost
$3.75

Limits

Output tokens
65,536 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare Gemini 3.8 Flash pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemini 3.8 Flash

Kilo Gateway

CoverageBenchmark

DataCamp's deep-dive on Gemini 3.8 Flash, published three days after the September 2, 2026 launch, reports concrete benchmark gains over the prior Flash: Terminal-Bench 2.1 jumps to 90.8% from 81.6% for Gemini 3.7 Flash, HLE-Verified reaches 54.9%, and DeepSWE v1.1 leadership is reaffirmed for long-horizon coding tasks The piece uniquely details the Gemini 3.8 Flash Cyber variant, a cybersecurity-tuned model achieving above 70% real-world vulnerability discovery and sitting on the CWE-Bench Pareto frontier for patching. Crucially, Cyber is not generally available: access is restricted to trusted defenders through Google's Fairwind Pr

Kilo Gateway

CoverageRelease Notes

Google announced Gemini 3.8 Flash on September 2, 2026 — its third Flash-tier release in six weeks — positioning it as the company's strongest reasoning and coding model in the Flash line. The model ships in two variants: a standard "workhorse" Flash aimed at agentic tasks and software development, and Gemini 3.8 Flash For developers, API access is available at an introductory rate of $0.75 per million input tokens and $3.75 per million output tokens through the end of 2026, after which regular pricing doubles to $1.50 and $7.50. The Ars Technica report frames the rapid Flash release cadence — and a reportedly delayed Gemini 3.5 Pro

Videos about Gemini 3.8 Flash

More models around Gemini 3.8 Flash