Currently listed through these providers:
Model details
Gemini 2.5 Flash-Lite
Gemini 2.5 Flash-Lite sits as the quick, efficient sibling inside the Gemini 2.5 family, engineered for ultra-low latency and cost efficiency rather than top-tier reasoning depth. Independent reporting frames it as roughly 1.5 times faster than its predecessor Gemini 2.0 Flash, with measured throughput of about 392.8 tokens per second and a sub-300-millisecond time-to-first-token that lets responses begin streaming before a user finishes typing. To preserve that speed profile, multi-pass "thinking" is switched off by default; developers who need heavier reasoning can opt back in through a Reasoning API parameter, trading extra tokens for stronger outputs on tasks like math or code generation.
Practically, the model fits high-volume, latency-sensitive workloads where premium-tier reasoning would be overkill or too expensive: large-scale chatbots, classification pipelines, bulk document processing, and long-context retrieval across books, PDFs, or sizable codebases. It accepts multimodal inputs through a single API while returning text only, keeping downstream integration simple, and the reasoning toggle plus aggressive cost economics make it well suited as a routing or filtering layer that escalates only the hardest queries to larger Gemini variants.
Quick Info
Powered by- Provider
- Merge Gateway
- Model key
- google/gemini-2.5-flash-lite
- Release date
- Jun 17, 2025
- Last updated
- Jun 17, 2025
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.10
- Output token cost
- $0.40
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,000,000 tokens
Transparent token rates
Compare Gemini 2.5 Flash-Lite pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Gemini 2.5 Flash-Lite
Videos about Gemini 2.5 Flash-Lite
More models around Gemini 2.5 Flash-Lite
This exact model name is also listed by 20 other providers.