Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
SAP AI Core logo

Model details

gemini-2.5-flash-lite

Gemini 2.5 Flash-Lite serves as the high-efficiency member of the Gemini 2.5 family, engineered specifically to maximize intelligence per dollar. Designed for developers and enterprise builders, the model prioritizes rapid response times and high-throughput performance, making it suitable for tasks that require immediate output, such as intelligent routing, translation, and large-scale classification. Its architecture supports a massive 1 million-token context window, allowing users to process entire books, extensive codebases, or long documents without the need for manual chunking.

Built to push the boundaries of speed, the model features native reasoning capabilities that can be toggled on to improve accuracy in math and code generation tasks. This flexibility allows users to balance performance and cost dynamically based on the complexity of their specific use case. As a production-ready tool, it excels in high-scale operations where efficiency is paramount, offering a streamlined path for integrating advanced AI into mission-critical workflows. Its design lineage focuses on maintaining a balance between rapid, low-latency performance and the reasoning depth required for demanding enterprise applications.

SAP AI Coregemini-2.5-flash-litegemini-flash-lite

Quick Info

Powered by
Provider
SAP AI Core
Model key
gemini-2.5-flash-lite
Release date
Jun 17, 2025
Last updated
Jun 17, 2025
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.10
Output token cost
$0.40

Limits

Output tokens
65,536 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare gemini-2.5-flash-lite pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about gemini-2.5-flash-lite

SAP AI Core

CoverageComparison

A data-driven comparison of Qwen3.5-Flash and Gemini 2.5 Flash-Lite - two models at the exact same $0.10/$0.40 per million token price point with 1M context windows but very different performance profiles.

SAP AI Core

CoveragePreview

Google is pulling Gemini 2.5 Flash-Lite Preview from AI Studio on March 31. The replacement, Gemini 3.1 Flash-Lite Preview, costs significantly more per token.

SAP AI Core

CoverageComparison

As of March 20, 2026, Gemini 2.5 Flash-Lite is still the better default if your main goal is the lowest stable token cost, while Gemini 3.1 Flash-Lite is the stronger successor lane if you can justify a much higher price for better quality and an eventual migration path. This guide explains when to stay, when to switch

SAP AI Core

Coverage

The new model aims to address a significant challenge enterprise developers face by providing levels of thinking to better match the task at hand.

SAP AI Core

CoveragePreview

The Latest Gemini 2.5 Flash-Lite Preview is Now the Fastest Proprietary Model (External Tests) and 50% Fewer Output Tokens.

SAP AI Core

Coverage

Google has been on a roll lately with its Gemini lineup of large language models, and now the company is expanding the 2.5 family with a new addition., Google has been on a roll lately with its Gemini lineup of large language models, and now the company is expanding the 2.5 family with a new addition.

Videos about gemini-2.5-flash-lite

More models around gemini-2.5-flash-lite