Currently listed through these providers:
Model details
Gemini 3.5 Flash
Gemini 3.5 Flash is positioned as the next step in the Gemini 3 line of capable, natively multimodal reasoning models, extending a Flash-tier foundation aimed at balancing quality, cost, and latency. According to the DeepMind model card, the model builds on the Gemini 3 Flash reasoning base and introduces thinking-level controls that let developers tune the trade-off between response depth, speed, and spend for a given workload. The published model card dates to May 2026, giving teams a current reference point for behavior and intended usage on Neon's platform.
In practical terms, this lineage matters because a Flash-tier reasoning model with adjustable thinking budgets tends to fit interactive applications where latency and budget must be controlled without giving up the ability to ask for deeper analysis on demand. Multimodal input handling means the same model can be used across text, image, audio, video, and document workflows, simplifying stack decisions for product teams. For developers choosing where Gemini 3.5 Flash fits, it is best understood as a versatile workhorse for production assistants, document and media understanding pipelines, and agentic flows that occasionally need more deliberate reasoning rather than a specialty model aimed at a single narrow task.
Quick Info
Powered by- Provider
- Neon
- Model key
- gemini-3-5-flash
- Release date
- May 19, 2026
- Last updated
- May 19, 2026
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.50
- Output token cost
- $9.00
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,048,576 tokens
Latest news about Gemini 3.5 Flash
Videos about Gemini 3.5 Flash
More models around Gemini 3.5 Flash
This exact model name is also listed by 21 other providers.