Sulat.com
AI models
Merge Gateway logo

Model details

Gemini 3.8 Flash

Gemini 3.8 Flash is the latest Flash-tier entry in the Gemini 3 family, succeeding Gemini 3.7 Flash and continuing Google's rapid iteration cadence on this line of models. According to Google DeepMind's product page, the model is positioned as "our most intelligent workhorse model yet for coding and agents," with a stated focus on complex agentic tasks at scale. It builds on its predecessor's customizable effort levels, letting developers tune the trade-off between quality, cost, and latency for each request rather than treating throughput as fixed.

In practice, Gemini 3.8 Flash is being marketed for long-horizon software engineering and autonomous agent workflows where sustained reasoning and tool use matter more than peak single-shot intelligence. Independent reporting from Ars Technica and DataCamp shows notable benchmark gains in coding and tool use over the previous Flash generation, including a jump from 81.6% to 90.8% on Terminal-Bench 2.1 and the top spot on the DeepSWE v1.1 long-horizon coding leaderboard, while broad-knowledge evaluations like Humanity's Last Exam stayed roughly flat at 45.4%. The result is a model that fits teams building production coding assistants and agent pipelines who want Flash-level latency and cost paired with stronger software-engineering competence than the prior release.

Merge Gatewaygoogle/gemini-3.8-flashgemini-flash

Quick Info

Powered by
Provider
Merge Gateway
Model key
google/gemini-3.8-flash
Release date
Sep 2, 2026
Last updated
Sep 2, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.75
Output token cost
$3.75

Limits

Output tokens
65,536 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare Gemini 3.8 Flash pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemini 3.8 Flash

Merge Gateway

CoverageBenchmark

DataCamp's technical blog details the September 2, 2026 release of Gemini 3.8 Flash, Google's third Flash model in six weeks. Key benchmarks include 90.8% on Terminal-Bench 2.1 (up from 81.6% for 3.7 Flash), 54.9% on HLE-Verified, and top placement on DeepSWE v1.1 for long-horizon coding. Pricing holds at $0.75 input a The post frames Gemini 3.8 Flash as Google's most intelligent Flash-tier model, engineered for long-horizon software engineering and autonomous agents, with a 1M-token context window and 64K max output. It notes uneven gains: coding and tool use jumped significantly while Humanity's Last Exam stayed flat at 45.4% versu

Merge Gateway

CoverageRelease Notes

Ars Technica's coverage of the September 2, 2026 launch confirms Gemini 3.8 Flash arrives as Google's third Flash release in six weeks, making a Gemini 3.5 Pro release increasingly unlikely. The standard Flash is positioned as a "workhorse" model for agentic tasks and software development, while Gemini 3.8 Flash Cyber Google's benchmark claims show Gemini 3.8 Flash topping the DeepSWE leaderboard for complex software engineering at lower cost, with marginal improvements over 3.7 Flash in most tests but larger coding gains. Computer use remains a struggle: while 3.8 Flash improved over 3.7 Flash on OSWorld-2.0, it still trails Claude

Videos about Gemini 3.8 Flash

More models around Gemini 3.8 Flash