Currently listed through these providers:
Model details
Gemini 3.5 Flash
Gemini 3.5 Flash was announced on May 19, 2026, during Google I/O, with Google framing it as delivering Pro-level reasoning at Flash-class latency. The model sits inside the Gemini 3.5 family and is built on the Gemini 3 Flash reasoning foundation, which introduces explicit thinking levels that let developers tune quality, cost, and latency for different workloads. Google is positioning this variant as a mid-tier option that can absorb agentic and coding tasks previously handled by the Pro tier, signaling an ambition to shift heavier production workloads onto a faster, more economical tier without giving up reasoning depth.
Independent evaluation has so far focused on whether that positioning holds up in practice. Appwrite's deep-dive compared the model using three sources: Google's published model card, the Artificial Analysis leaderboard entry, and Appwrite Arena, an open-source benchmark covering 191 questions across nine Appwrite service categories. That combination of vendor documentation, third-party leaderboard data, and a reproducible service-oriented benchmark gives developers a reasonable basis to test whether the model's reasoning and agentic coding claims translate into real production environments, particularly for retrieval, orchestration, and multi-step tool-driven workflows.
Quick Info
Powered by- Provider
- Requesty
- Model key
- gemini-3.5-flash
- Release date
- May 19, 2026
- Last updated
- May 19, 2026
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.50
- Output token cost
- $9.00
Limits
- Output tokens
- 65,535 tokens
- Context window
- 1,048,576 tokens
Latest news about Gemini 3.5 Flash
Videos about Gemini 3.5 Flash
Recent tweets and retweets from Requesty
More models around Gemini 3.5 Flash
This exact model name is also listed by 21 other providers.