Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Vertex logo

Model details

Gemini 3.1 Flash Lite Preview

Gemini 3.1 Flash-Lite Preview is positioned as a lean entry in Google's Gemini family, documented on the official Gemini API site under the model id gemini-3.1-flash-lite-preview. The Preview label signals an experimental release intended for early evaluation and integration testing rather than long-term production anchoring. As a Flash-Lite tier variant, the design emphasis is on cost-efficient inference while retaining the broader Gemini multimodal capability set, making it suitable for high-volume, latency-sensitive workloads where a heavier Pro or standard Flash model would be overkill.

The model accepts a broad multimodal input palette covering text, image, video, audio, and PDF, while producing text-only outputs, which is typical for reasoning-and-response style assistants in this tier. Its placement in the Gemini 3.x lineup suggests continuity with the family's unified multimodal training approach, where shared embedding and instruction tuning allow a single checkpoint to serve diverse input types. The surrounding documentation context also points forward to a maturing Gemini ecosystem, with newer siblings like Gemini 3.8 Flash referenced alongside, reinforcing that Flash-Lite Preview is one node in an actively evolving model family rather than a standalone endpoint.

Vertexgemini-3.1-flash-lite-previewgemini-flash-litedeprecated

Quick Info

Powered by
Provider
Vertex
Model key
gemini-3.1-flash-lite-preview
Release date
Mar 3, 2026
Last updated
Mar 3, 2026
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.25
Output token cost
$1.50

Limits

Output tokens
65,536 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare Gemini 3.1 Flash Lite Preview pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemini 3.1 Flash Lite Preview

Vertex

Coverage

Google has introduced Gemini 3.1 Flash-Lite, its fastest and most cost-effective AI model yet, offering enhanced performance for a variety of tasks.

Vertex

Coverage

Google unveils Gemini 3.1 Flash Lite, its fastest and most cost-efficient AI model for developers, offering scalable deployment via Gemini API and Vertex AI.

Vivgrid

CoverageRelease Notes

The Opper AI Google model release tracker lists Gemini 3.1 Flash Lite Preview among dated Google releases, showing it launched on 3 Mar 2026 with a 1M token context window and listed token pricing of $0.25 input / $1.50 output, alongside an Intelligence score of 16. The entry is part of a broader timeline that places t Beyond the dated launch entry, the Opper page does not provide capability details, benchmarks, or technical descriptions specific to Gemini 3.1 Flash Lite Preview; it functions as a release index rather than substantive news. No claim in the excerpt asserts any Vivgrid-specific gateway or serving behavior. The page is

Videos about Gemini 3.1 Flash Lite Preview

More models around Gemini 3.1 Flash Lite Preview