Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
LLM Gateway logo

Model details

Gemini 3.1 Flash Lite (Google Vertex AI)

Gemini 3.1 Flash Lite targets the lightweight end of Google's Gemini family, positioned as a cost- and latency-focused variant for high-frequency, agent-driven traffic. Rather than emphasizing frontier reasoning, the design centers on simple data extraction, routing, and other utility calls where response time and per-token spend matter more than raw capability. It accepts a broad mix of inputs — text, images, video, audio, and PDFs — while returning text, which makes it flexible for pipelines that need to interpret mixed documents or media snippets without committing to a larger, slower model.

In practical terms, the model pairs a one-million-token context window with a 65,536-token output ceiling, giving it room for substantial documents or long agent histories while keeping generation bounded and predictable. It supports tool calling, structured output, and reasoning steps alongside streaming, JSON mode, batch processing, and fine-tuning, which makes it straightforward to drop into existing agent frameworks that expect function calls and typed responses. The combination of a very low input price and a modest output price further reinforces its fit for high-volume workloads, where many short calls can be cheaper than relying on a larger Gemini tier without sacrificing multimodal coverage.

LLM Gatewaygoogle-vertex/gemini-3.1-flash-litegemini-flash-lite

Quick Info

Powered by
Provider
LLM Gateway
Model key
google-vertex/gemini-3.1-flash-lite
Release date
May 7, 2026
Last updated
May 7, 2026
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.25
Output token cost
$1.50

Limits

Output tokens
65,536 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare Gemini 3.1 Flash Lite (Google Vertex AI) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemini 3.1 Flash Lite (Google Vertex AI)

No articles yet. Fetch the latest news to show it here.

Videos about Gemini 3.1 Flash Lite (Google Vertex AI)

More models around Gemini 3.1 Flash Lite (Google Vertex AI)