Currently listed through these providers:
Model details
Gemini 3.1 Flash Lite Preview
Gemini 3.1 Flash Lite Preview positions itself as the efficiency-focused entry point within Google's Gemini 3.1 generation, designed specifically for budget-constrained, high-volume workloads. The model brings a notable advancement over its predecessor, Gemini 2.5 Flash Lite, with improvements spanning audio input and automatic speech recognition, RAG snippet ranking, translation accuracy, data extraction reliability, and code completion quality. What sets this model apart is its support for four configurable thinking levels—minimal, low, medium, and high—allowing developers to fine-tune the cost-performance trade-off depending on task complexity. This flexibility makes it particularly well-suited for applications where both speed and resource efficiency matter, such as real-time data processing pipelines, customer-facing chatbots, and automated content moderation systems.
The practical performance gains are backed by benchmark evidence showing strong results on reasoning and knowledge tasks, with particular strength in coding applications. Benchmark scores indicate solid performance on the GPQA Diamond evaluation for domain knowledge, while coding capabilities show measurable improvement over earlier iterations. Being priced at roughly half the cost of the full Gemini 3 Flash model, it strikes a balance between capability and affordability that appeals to teams scaling AI features across large user bases. The model maintains full multimodal input support for text, images, video, audio, and PDF documents, making it versatile for enterprise workflows that require processing diverse document types. This combination of incremental quality improvements, configurable reasoning behavior, and accessible pricing positions it as a practical choice for developers building production AI features at scale.
Quick Info
Powered by- Provider
- Merge Gateway
- Model key
- google/gemini-3.1-flash-lite-preview
- Release date
- Mar 3, 2026
- Last updated
- Mar 3, 2026
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.25
- Output token cost
- $1.50
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,048,576 tokens
Transparent token rates
Compare Gemini 3.1 Flash Lite Preview pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Gemini 3.1 Flash Lite Preview
Videos about Gemini 3.1 Flash Lite Preview
More models around Gemini 3.1 Flash Lite Preview
This exact model name is also listed by 9 other providers.