Currently listed through these providers:
Model details
Gemini 2.5 Flash
Gemini the listed price Flash represents Google's push to make advanced AI accessible for production workloads where cost efficiency matters alongside capability. Designed as a non-reasoning model optimized for speed, it prioritizes rapid output generation suitable for high-volume applications like classification, translation, and intelligent routing tasks. The architecture balances above-average intelligence scores with the reliability and scalability that enterprise deployments demand, offering organizations a stable foundation for mission-critical AI solutions.
Following its transition to general availability, this model underwent refinement to deliver consistent performance across demanding workloads. Analysis indicates notably fast output generation paired with robust handling of multimodal inputs, though the system tends toward more verbose responses. This combination positions it well for organizations building sophisticated, customized AI applications that require both responsiveness and the flexibility to handle diverse input types without sacrificing production-ready stability.
Quick Info
Powered by- Provider
- NEAR AI Cloud
- Model key
- google/gemini-2.5-flash
- Release date
- Jun 17, 2025
- Last updated
- Jun 17, 2025
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.30
- Output token cost
- $2.50
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,048,576 tokens