Currently listed through these providers:
Model details
Gemini Omni Flash Preview
Gemini Omni Flash Preview anchors the launch of Google's Omni family, positioning the system as the video counterpart to the brand's established image-generation work and bringing generative video into the same conversational workflow that the Gemini line is known for. Rather than behaving as a one-shot renderer, it is built around a stateful editing loop: a short clip with synchronized audio is produced from text, image, or video references, and the scene can then be refined turn by turn in natural language while leaving the portions of the clip the user did not mention untouched. This stateful-preservation design is what makes the model useful for iterative directing, where creators adjust small details across many rounds without rebuilding the surrounding footage each time.
In practice, the model fits creators and product teams who want to sketch, prototype, or produce short video assets directly from multimodal prompts rather than assembling them in a traditional editor, and it is exposed to developers through familiar entry points including Google AI Studio and the Gemini API, as well as integrations on platforms like Vercel AI Gateway that expose a streaming interface. As an early entrant in Google's Omni lineup, it sets a baseline for what multimodal video generation can look like inside the Gemini ecosystem, with the preview tag signaling that the capability surface and quality bar are still evolving and that further siblings in the family are likely to build on the same conversational editing approach.
Quick Info
Powered by- Provider
- Model key
- gemini-omni-flash-preview
- Release date
- Jun 30, 2026
- Last updated
- Jun 30, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.50
- Output token cost
- $17.50
Limits
- Output tokens
- 65,536 tokens
- Context window
- 131,072 tokens