Currently listed through these providers:
Model details
Gemini 3.5 Flash Lite
Gemini 3.5 Flash-Lite belongs to Google’s 3.5-class Gemini family and is positioned for low-latency, high-throughput agentic work. Google DeepMind describes it as the family’s fastest and most cost-effective option, with an Artificial Analysis Index attribution of 350 output tokens per second; that figure should be understood as a third-party benchmark result rather than a guarantee for every workload.
This model is intended for varied tasks that benefit from quick responses, including coding, UI generation, and translation. Its practical fit is applications where throughput, responsiveness, and operating cost matter more than heavyweight processing, especially when an AI Studio entry point and dedicated API documentation can support evaluation and integration.
Quick Info
Powered by- Provider
- Abacus
- Model key
- gemini-3.5-flash-lite
- Release date
- Jul 21, 2026
- Last updated
- Jul 21, 2026
- Knowledge cutoff
- 2026-03
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.30
- Output token cost
- $2.50
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,048,576 tokens
Latest news about Gemini 3.5 Flash Lite
Videos about Gemini 3.5 Flash Lite
Recent tweets and retweets from Abacus
More models around Gemini 3.5 Flash Lite
This exact model name is also listed by 21 other providers.