Model details
Agnes 3.0 Flash
Agnes 3.0 Flash is Agnes AI's next-generation text model designed for agentic coding and tool-driven development workflows. According to the official documentation, it is built to handle the full execution path of real-world agent tasks, from understanding requirements and planning through tool invocation to final delivery. The model places deliberate emphasis on stability, instruction following, grounded execution, and output integrity when handling complex multi-step work, qualities that matter most when an agent must chain tools and reasoning across many turns.
With a 512K-token context window and a 65,536-token maximum output, Agnes 3.0 Flash is shaped for long-task context retention rather than short single-turn exchanges. It accepts text and image-URL inputs while producing text outputs, and is exposed through familiar Chat Completions, Responses, and Messages endpoints, which makes it straightforward to integrate into existing agent stacks. Third-party listing also highlights practical agent-oriented capabilities such as thinking, streaming, tool calling, structured outputs, prompt caching, and URL context, all of which reinforce its positioning as a dependable engine for developers building reliable, tool-using applications.
Quick Info
Powered by- Provider
- NanoGPT
- Model key
- agnes-3.0-flash
- Release date
- Sep 9, 2026
- Last updated
- Sep 9, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.05
- Output token cost
- $0.15
Limits
- Input tokens
- 524,288 tokens
- Output tokens
- 65,536 tokens
- Context window
- 524,288 tokens
Latest news about Agnes 3.0 Flash
No articles yet. Fetch the latest news to show it here.