Vertex
Google has introduced Gemini 3.1 Flash-Lite, its fastest and most cost-effective AI model yet, offering enhanced performance for a variety of tasks.
Model details
Gemini 3.1 Flash-Lite Preview is positioned as a lean entry in Google's Gemini family, documented on the official Gemini API site under the model id gemini-3.1-flash-lite-preview. The Preview label signals an experimental release intended for early evaluation and integration testing rather than long-term production anchoring. As a Flash-Lite tier variant, the design emphasis is on cost-efficient inference while retaining the broader Gemini multimodal capability set, making it suitable for high-volume, latency-sensitive workloads where a heavier Pro or standard Flash model would be overkill.
The model accepts a broad multimodal input palette covering text, image, video, audio, and PDF, while producing text-only outputs, which is typical for reasoning-and-response style assistants in this tier. Its placement in the Gemini 3.x lineup suggests continuity with the family's unified multimodal training approach, where shared embedding and instruction tuning allow a single checkpoint to serve diverse input types. The surrounding documentation context also points forward to a maturing Gemini ecosystem, with newer siblings like Gemini 3.8 Flash referenced alongside, reinforcing that Flash-Lite Preview is one node in an actively evolving model family rather than a standalone endpoint.
Transparent token rates
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Vertex
Google has introduced Gemini 3.1 Flash-Lite, its fastest and most cost-effective AI model yet, offering enhanced performance for a variety of tasks.
Vertex
Google unveils Gemini 3.1 Flash Lite, its fastest and most cost-efficient AI model for developers, offering scalable deployment via Gemini API and Vertex AI.
Vivgrid
The Opper AI Google model release tracker lists Gemini 3.1 Flash Lite Preview among dated Google releases, showing it launched on 3 Mar 2026 with a 1M token context window and listed token pricing of $0.25 input / $1.50 output, alongside an Intelligence score of 16. The entry is part of a broader timeline that places t Beyond the dated launch entry, the Opper page does not provide capability details, benchmarks, or technical descriptions specific to Gemini 3.1 Flash Lite Preview; it functions as a release index rather than substantive news. No claim in the excerpt asserts any Vivgrid-specific gateway or serving behavior. The page is
This exact model name is also listed by 10 other providers.