Currently listed through these providers:
Model details
Gemini 3.5 Flash (Vertex AI, EU)
Gemini 3.5 Flash is positioned as a fast model in the Gemini Flash family, balancing multimodal reasoning, tool use, and cost for real-world applications. Its closed-weight design places it among proprietary offerings, while the available evidence emphasizes practical performance rather than documented architectural or training-scale details.
The model is best suited to agentic workflows, sub-agent deployment, multi-step tasks, and long-running applications that benefit from reasoning combined with tool calling. Independent usage data from Requesty reports roughly 197 generated tokens per second in aggregate and a 967-millisecond best first-token time across the listed Google routes, supporting a fit for latency-sensitive, high-volume work. Provider routes can differ in speed and cost, so the EU-specific endpoint should be evaluated against the exact workload and route.
Quick Info
Powered by- Provider
- Eden AI
- Model key
- vertex/gemini-3.5-flash@eu
- Release date
- May 19, 2026
- Last updated
- May 19, 2026
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.50
- Output token cost
- $9.00
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,048,576 tokens
Transparent token rates
Compare Gemini 3.5 Flash (Vertex AI, EU) pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Gemini 3.5 Flash (Vertex AI, EU)
No articles yet. Fetch the latest news to show it here.