Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Requesty logo

Model details

Gemini 3.8 Flash (EU)

Gemini 3.8 Flash (EU) is a region-pinned European deployment of Google's Gemini 3.8 Flash model, surfaced through Requesty as a router entry that runs on Vertex AI infrastructure hosted in the EU. The endpoint is configured so that customer data is not retained and is not used for training, which makes it a practical choice for EU compliance-sensitive workloads where data residency and isolation matter. Because it is served through Vertex AI, teams can tap into Google's managed serving stack while still routing access and billing through Requesty's unified catalog.

For application design, the deployment offers a 1.0M-token context window with up to roughly 66K output tokens on the chat API type, which is well suited to long-context retrieval-augmented generation, multi-document summarization, and agentic pipelines that need to keep large transcripts or code repositories in scope at once. Through Requesty the EU endpoint is priced at a 50% discount off Vertex list rates, giving teams an economical way to run high-volume Flash-class inference with strong EU data-residency guarantees. The combination of a large context window, Flash-class economics, and EU pinning makes this listing a sensible default for production assistants and backend services that need to stay within European regulatory boundaries.

Requestygemini-3.8-flash@eugemini-flash

Quick Info

Powered by
Provider
Requesty
Model key
gemini-3.8-flash@eu
Release date
Sep 2, 2026
Last updated
Sep 2, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.825
Output token cost
$4.125

Limits

Output tokens
65,535 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare Gemini 3.8 Flash (EU) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemini 3.8 Flash (EU)

Requesty

Official sourceBenchmark

Requesty has added a region-pinned EU deployment of Google's Gemini 3.8 Flash to its router catalog, identified as `vertex/gemini-3.8-flash@eu` and added in September 2026. The endpoint runs on Vertex AI infrastructure served from the EU with no data retention and no use for training, making it suitable for workloads t The deployment offers a 1.0M-token context window with up to 66K tokens of output on the chat API type, and Requesty is pricing it at a 50% discount off Vertex list rates: $0.83 per 1M input, $4.13 per 1M output, and $0.08 per 1M cached input read (list: $1.65 / $8.25). Sample workload costs from the catalog include ro

Videos about Gemini 3.8 Flash (EU)

More models around Gemini 3.8 Flash (EU)