Currently listed through these providers:
Model details
GPT Chat Latest
GPT Chat Latest operates as a stable API alias that always resolves to the latest Instant chat model powering ChatGPT, designed specifically for developers who want frictionless access to ongoing improvements without updating their integrations. Rather than pinning to a fixed model version, this endpoint automatically routes to new Instant model updates as OpenAI releases them, making it a forward-looking choice for applications that benefit from incremental capability gains. Its support for function calling, structured outputs, and reasoning mode positions it as a versatile option for complex workflows like agent orchestration, automated coding tasks, and multi-step reasoning chains.
The model demonstrates strong real-world adoption through production applications, with Hermes Agent and Kilo Code among the high-traffic users processing hundreds of millions of tokens weekly, indicating reliable performance at scale. Performance metrics show throughput around 70 tokens per second with end-to-end latency averaging 3.29 seconds, and error rates for tool calling and structured output both stay below half a percent, reflecting a well-tuned inference pipeline. Being proprietary rather than open-weight, it benefits from OpenAI's centralized training and alignment pipeline while offering the convenience of automatic model rotation behind a consistent API slug.
Quick Info
Powered by- Provider
- OpenRouter
- Model key
- openai/gpt-chat-latest
- Release date
- May 5, 2026
- Last updated
- May 5, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $5.00
- Output token cost
- $30.00
Limits
- Output tokens
- 128,000 tokens
- Context window
- 400,000 tokens
Transparent token rates
Compare GPT Chat Latest pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GPT Chat Latest
No articles yet. Fetch the latest news to show it here.
