Poe
Google is pulling Gemini 2.5 Flash-Lite Preview from AI Studio on March 31. The replacement, Gemini 3.1 Flash-Lite Preview, costs significantly more per token.
Model details
Gemini 2.5 Flash belongs to Google's family of thinking models, designed to reason through their thoughts before generating responses, with developers able to control a thinking budget that determines how much deliberation occurs prior to output. Google formally documented this model line in a June 2025 update announcing that Gemini 2.5 Flash had reached general availability and stability. The model is positioned within Google's Gemini Enterprise Agent Platform, where it serves as a documented, supported offering for enterprise and developer use cases.
As of early 2026, independent reporting indicates Gemini 2.5 Flash has been effectively superseded in Google's lineup by Gemini 3.1 Flash-Lite, which Google positions as faster and more cost-efficient at scale. The Gemini 2.5 Flash name also encompasses distinct sub-variants, including a separate image-editing model known as nano-banana, which has attracted viral attention for improvements in character consistency and multi-turn photo editing. Third-party benchmark comparisons continue to treat Gemini Flash as a recognized baseline, with entrants like Sarvam's 105B mixture-of-experts model claiming to outperform it on key evaluations, underscoring Gemini 2.5 Flash's continued relevance as a reference point in the public benchmark landscape.
Transparent token rates
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Poe
Google is pulling Gemini 2.5 Flash-Lite Preview from AI Studio on March 31. The replacement, Gemini 3.1 Flash-Lite Preview, costs significantly more per token.
Poe
Google has launched Gemini 3.1 Flash-Lite, its fastest and most cost-efficient AI model, at $0.25 per million tokens and 2.5x faster than Gemini 2.5 Flash.
Poe
Startup unveils two mixture-of-experts LLMs trained from scratch as part of India’s sovereign AI push
Poe
Google has unveiled Gemini 2.5 Flash Image, the viral ‘nano-banana’ AI model, bringing major upgrades to photo editing with better character consistency, multi-turn edits, and built-in watermarks.
Poe
Google's official Gemini Apps release notes (entry dated 2026.07.21) announce "Gemini 3.6 Flash," described as an upgraded Flash model for everyday tasks, rolling out to all Gemini app users globally and accessible via the "3.6 Flash" option in the model's drop-down. Google frames 3.6 Flash as delivering speed and qual The same release-note entry is directly relevant context for third-party aggregators serving the older 2.5 Flash SKU: it implies a generational shift in Google's Flash model family that does not yet announce a specific Gemini API or Vertex AI deprecation for 2.5 Flash. The page does not address Poe's `google/gemini-2.5