Sulat.com
AI models
Vertex logo

Model details

Gemini 2.5 Flash-Lite

Gemini 2.5 Flash-Lite is engineered as a fast, cost-efficient model optimized for high-scale, cost-sensitive operations. Designed as the lightweight sibling within the Flash family, it prioritizes throughput and affordability without sacrificing core capabilities. This model is well-suited for tasks like classification, intelligent routing, and batch translation, making it a practical choice for production pipelines that need to process large volumes of requests without ballooning costs.

The September 2025 update brought three focused improvements: better instruction following for complex system prompts, reduced verbosity to trim token usage, and stronger multimodal capabilities including more accurate audio transcription and improved image understanding. Agentic tool use also saw meaningful gains—SWE-Bench Verified scores rose from 48.9% to 54%, a 5% improvement that signals better performance in multi-step, tool-driven applications. With thinking enabled, the model achieves higher quality outputs while consuming fewer tokens, reducing both latency and cost for real-world deployments.

Vertexgemini-2.5-flash-litegemini-flash-lite

Quick Info

Powered by
Provider
Vertex
Model key
gemini-2.5-flash-lite
Release date
Jun 17, 2025
Last updated
Jun 17, 2025
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.10
Output token cost
$0.40

Limits

Output tokens
65,536 tokens
Context window
1,048,576 tokens

Latest news about Gemini 2.5 Flash-Lite

No articles yet. Fetch the latest news to show it here.

Videos about Gemini 2.5 Flash-Lite

More models around Gemini 2.5 Flash-Lite