Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Eden AI logo

Model details

Gemini 3.8 Flash (Vertex AI)

Gemini 3.8 Flash is positioned within Google's Gemini family as a workhorse model focused on coding and agent-driven workflows. Google introduced it on September 2, 2026 alongside a sibling variant called Gemini 3.8 Flash Cyber, framing the release as delivering next-generation intelligence for agentic workflows and cybersecurity use cases. The DeepMind product page reinforces this framing by describing the model as the most intelligent workhorse model yet for coding and agents, signaling that the Flash tier is intended to balance capability with practical deployment scale rather than chasing maximum reasoning depth at any cost.

Practically, the model is aimed at developers and teams building complex agentic pipelines that require reliable tool use, structured reasoning, and steady throughput. DeepMind highlights the model as best for tackling complex agentic tasks at scale, and the product page exposes dedicated sections for capabilities, hands-on exploration, showcase examples, performance details, and model information, suggesting a mature documentation surface for production integration. Access is offered both to end users through the Gemini consumer app and to builders through Google AI Studio, making it a natural fit for engineering teams that want to prototype agentic workflows quickly while retaining a path to scaled deployment.

Eden AIvertex/gemini-3.8-flashgemini-flash

Quick Info

Powered by
Provider
Eden AI
Model key
vertex/gemini-3.8-flash
Release date
Sep 2, 2026
Last updated
Sep 2, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.75
Output token cost
$3.75

Limits

Output tokens
65,536 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare Gemini 3.8 Flash (Vertex AI) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemini 3.8 Flash (Vertex AI)

Eden AI

CoverageBenchmark

Google has shipped three Flash releases in about six weeks — Gemini 3.6 Flash on 21 July 2026, 3.7 Flash on 13 August 2026, and 3.8 Flash on 2 September 2026 — with Google calling 3.8 Flash its "most intelligent Flash model". The published spec line is identical to the previous Flash model, with the same 1,048,576 inpu Gemini 3.8 Flash adds low, medium and high thinking levels (up from minimal/low/medium/high on 3.6 and the low/medium/high set on 3.7) and Google benchmarks it against Claude Opus 5, Claude Sonnet 5 and the GPT-5.6 line rather than the newest OpenAI flagship. ComputingForGeeks ran calls against gemini-3.8-flash in Sept

Eden AI

CoverageBenchmark

Gemini 3.8 Flash is GA as the direct successor to Gemini 3.7 Flash, published with the model ID gemini-3.8-flash and a one-million-token context window (1,048,576 input tokens) with a maximum output of 65,536 tokens. Input modalities are text, image, video, audio and PDF, with text output, and supported tools include c Google positions the model as a meaningful upgrade for long-horizon coding, tool-using agents, finance and legal workflows, and scientific research tasks, while keeping the temporary $0.75 per 1M input and $3.75 per 1M output token prices. Kingy AI's six-call low-thinking API smoke check found that at high effort the m

Eden AI

CoverageRelease Notes

Google released Gemini 3.8 Flash on September 2, 2026, marking its third Flash-tier model launch in just six weeks. The release comes in two variants: a standard "workhorse" Flash model suited to agentic tasks and software development, and Gemini 3.8 Flash Cyber, which is tuned for vulnerability detection and mitigatio For developers, API access is priced at an introductory rate of $0.75 per million input tokens and $3.75 per million output tokens through the end of 2026, doubling to $1.50/$7.50 thereafter. Google published benchmark numbers showing marginal gains over Gemini 3.7 Flash on most tests but larger improvements in coding

Eden AI

CoverageBenchmark

Google released Gemini 3.8 Flash on September 2, 2026, three weeks after Gemini 3.7 Flash and one day after Anthropic shipped Claude Fable 5.1, with the pitch of frontier-class results on coding, finance, and legal agent benchmarks at a fraction of rival prices. Per Google's official model card, the model is built for Specifications include a 1-million-token context window, 64,000-token maximum output, multimodal inputs (text, images, audio, video), and customizable effort levels that trade quality against cost and latency. Knowledge cutoffs are mixed at March 2026 for some domains and January 2025 for others, with introductory pric

Eden AI

CoverageBenchmark

Third-party benchmark aggregator BenchLM tracks Gemini 3.8 Flash as a proprietary Google DeepMind release that went live on September 2, 2026 with a 1-million-token context window. The model scores 78.4 out of 100 overall and ranks 6th out of 232 tracked models, with its strongest eligible category being Coding at rank Speed and pricing metrics from BenchLM show 327 tokens per second output throughput with a 10.75-second time-to-first-token. API pricing is reported at $0.75 per million input tokens and $3.75 per million output tokens, with cached input priced at $0.075 per million and a blended rate of about $2.25 per million. The da

Eden AI

CoverageAnalysis

Google DeepMind released Gemini 3.8 Flash on September 2, 2026, as its fourth Flash model in under four months, positioning it as the most intelligent Flash-tier model to date. Artificial Analysis reports that the model scores 59 on the Intelligence Index with high reasoning, up three points from Gemini 3.7 Flash (high The three-point intelligence gain is driven primarily by stronger performance on agentic evaluations, including τ³-Banking (tool use), Terminal-Bench v2.1 (coding) and GDPval-AA v2 (real-world tasks), with the largest improvement on τ³-Banking at plus 12 points to 45 percent. Gemini 3.8 Flash inherits Gemini 3.7 Flash'

Eden AI

CoverageBenchmark

Google released Gemini 3.8 Flash on September 2, 2026 as its most intelligent Flash-tier model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows, sitting below the Pro tier while priced at a fraction of the cost. DataCamp reports it scoring 90.8 percent on Terminal-B Pricing holds at $0.75 input and $3.75 output per 1M tokens through December 31, 2026, then reverts to $1.50 and $7.50. Alongside the base model, Google introduced Gemini 3.8 Flash Cyber, a variant tuned for cybersecurity work with a real-world vulnerability discovery rate above 70 percent and a position on the CWE-Ben

Videos about Gemini 3.8 Flash (Vertex AI)

More models around Gemini 3.8 Flash (Vertex AI)