Sulat.com
AI models
Merge Gateway logo

Model details

Gemini 3.1 Flash Lite Preview

Gemini 3.1 Flash Lite Preview positions itself as the efficiency-focused entry point within Google's Gemini 3.1 generation, designed specifically for budget-constrained, high-volume workloads. The model brings a notable advancement over its predecessor, Gemini 2.5 Flash Lite, with improvements spanning audio input and automatic speech recognition, RAG snippet ranking, translation accuracy, data extraction reliability, and code completion quality. What sets this model apart is its support for four configurable thinking levels—minimal, low, medium, and high—allowing developers to fine-tune the cost-performance trade-off depending on task complexity. This flexibility makes it particularly well-suited for applications where both speed and resource efficiency matter, such as real-time data processing pipelines, customer-facing chatbots, and automated content moderation systems.

The practical performance gains are backed by benchmark evidence showing strong results on reasoning and knowledge tasks, with particular strength in coding applications. Benchmark scores indicate solid performance on the GPQA Diamond evaluation for domain knowledge, while coding capabilities show measurable improvement over earlier iterations. Being priced at roughly half the cost of the full Gemini 3 Flash model, it strikes a balance between capability and affordability that appeals to teams scaling AI features across large user bases. The model maintains full multimodal input support for text, images, video, audio, and PDF documents, making it versatile for enterprise workflows that require processing diverse document types. This combination of incremental quality improvements, configurable reasoning behavior, and accessible pricing positions it as a practical choice for developers building production AI features at scale.

Merge Gatewaygoogle/gemini-3.1-flash-lite-previewgemini-flash-lite

Quick Info

Powered by
Provider
Merge Gateway
Model key
google/gemini-3.1-flash-lite-preview
Release date
Mar 3, 2026
Last updated
Mar 3, 2026
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.25
Output token cost
$1.50

Limits

Output tokens
65,536 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare Gemini 3.1 Flash Lite Preview pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemini 3.1 Flash Lite Preview

Videos about Gemini 3.1 Flash Lite Preview

More models around Gemini 3.1 Flash Lite Preview