Currently listed through these providers:
Model details
Gemini 2.5 Flash-Lite
Gemini 2.5 Flash-Lite is engineered as a fast, cost-efficient model optimized for high-scale, cost-sensitive operations. Designed as the lightweight sibling within the Flash family, it prioritizes throughput and affordability without sacrificing core capabilities. This model is well-suited for tasks like classification, intelligent routing, and batch translation, making it a practical choice for production pipelines that need to process large volumes of requests without ballooning costs.
The September 2025 update brought three focused improvements: better instruction following for complex system prompts, reduced verbosity to trim token usage, and stronger multimodal capabilities including more accurate audio transcription and improved image understanding. Agentic tool use also saw meaningful gains—SWE-Bench Verified scores rose from 48.9% to 54%, a 5% improvement that signals better performance in multi-step, tool-driven applications. With thinking enabled, the model achieves higher quality outputs while consuming fewer tokens, reducing both latency and cost for real-world deployments.
Quick Info
Powered by- Provider
- Vertex
- Model key
- gemini-2.5-flash-lite
- Release date
- Jun 17, 2025
- Last updated
- Jun 17, 2025
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.10
- Output token cost
- $0.40
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,048,576 tokens
Latest news about Gemini 2.5 Flash-Lite
No articles yet. Fetch the latest news to show it here.
Videos about Gemini 2.5 Flash-Lite
More models around Gemini 2.5 Flash-Lite
This exact model name is also listed by 19 other providers.