Currently listed through these providers:
Model details
Gemini 3.5 Flash
Gemini 3.5 Flash sits in Google's Gemini Flash tier, a line aimed at low-latency, cost-efficient inference for production applications. A Google Cloud documentation page titled "Gemini 3.5 Flash | Gemini Enterprise Agent Platform" exists, confirming the model is positioned for enterprise agent platforms rather than purely conversational use, and a MindStudio comparison published the day after catalog launch positions it against Gemini 3.1 Pro, signaling it is treated by third parties as a viable replacement for heavier Pro-tier deployments where speed matters more than peak reasoning depth. The Flash designation historically indicates a distilled or efficiency-tuned variant of the broader Gemini family, optimized for high-throughput tasks such as retrieval-augmented generation, tool orchestration, and structured data extraction inside agent pipelines.
OfoxAI exposes the model through a unified API gateway that offers OpenAI-compatible, Anthropic-native, and Gemini-native protocols under a single key, with the Gemini native path routed at the documented v1beta endpoint. This means developers can adopt Gemini 3.5 Flash using either standard OpenAI-style chat completions or Google's native generateContent format, which is useful for teams migrating multimodal workloads or building agent stacks that need first-class support for attachments, reasoning, tool calling, temperature control, and structured output. The combination of a large context window, multimodal inputs spanning text, image, video, audio, and PDF, and gateway-level routing makes this model a practical fit for assistants that must ingest rich documents while keeping per-token cost low at scale.
Quick Info
Powered by- Provider
- Ofox
- Model key
- google/gemini-3.5-flash
- Release date
- May 19, 2026
- Last updated
- May 19, 2026
- Knowledge cutoff
- 2025-01
- AI SDK package
@ai-sdk/google- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.50
- Output token cost
- $9.00
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,048,576 tokens
Latest news about Gemini 3.5 Flash
Videos about Gemini 3.5 Flash
More models around Gemini 3.5 Flash
This exact model name is also listed by 21 other providers.