Currently listed through these providers:
Model details
Gemini 3 Flash Preview
Gemini the listed price Flash Preview sits in the Gemini Flash family as a lightweight thinking model that tries to bring near-Pro reasoning and tool use into latency-sensitive applications. The listing frames it as a "high speed, high value" model purpose-built for agentic workflows, multi-turn chat, and coding assistance, where it can carry its own inside long-running agent loops and collaborative coding sessions while still keeping interactive response times tight. Compared to the prior Gemini 2.5 Flash, the provider describes it as a broad quality uplift across reasoning, multimodal understanding, and reliability, positioning it as a sensible default for developers who want frontier-style reasoning and tool-calling behavior without paying for the largest Gemini variants.
In practical terms, the model is shaped for developers wiring it into applications rather than for one-off chat prompts. It combines a 1M-token context window with multimodal intake across text, images, audio, video, and PDFs, while emitting text-only responses, so a single endpoint can digest long agent transcripts alongside rich media. Configurable reasoning levels (minimal, low, medium, high) let callers dial the depth of step-by-step thinking up or down depending on how much latency budget they have, and structured output plus tool use make it straightforward to drop into automated pipelines where the model's replies must be parsed by downstream code. That blend of long context, optionally deeper reasoning, and reliable tool/structured output is what makes it a natural fit for agentic products, coding copilots, and retrieval-heavy assistants that need to stay responsive.
Quick Info
Powered by- Provider
- OrcaRouter
- Model key
- google/gemini-3-flash-preview
- Release date
- Dec 17, 2025
- Last updated
- Dec 17, 2025
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.50
- Output token cost
- $3.00
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,048,576 tokens