Currently listed through these providers:
Model details
GPT-5 Chat Latest
GPT-5 Chat Latest is documented by AI/ML API as the non-reasoning variant of GPT-5, distinguishing it from sibling endpoints that emphasize chain-of-thought or extended deliberation. Its place in OpenAI's model catalog, under the dedicated "GPT-5 Chat Model" reference page, situates it alongside general text generation, code generation, model selection, and structured output guides, signaling a focus on everyday chat and drafting rather than specialized analytical tasks. Within AI/ML API's reference, the model is exposed under the identifier openai/gpt-5-chat-latest with a hosted Playground entry, making it straightforward to route conversational prompts through that gateway without modifying upstream tooling.
In practical terms, GPT-5 Chat Latest fits teams that want fluent, instruction-following dialogue backed by the GPT-5 family but without the latency or cost overhead of a reasoning-tuned endpoint. Its position inside the broader OpenAI documentation hierarchy, where text generation, code generation, and structured outputs share the same Models section, suggests it inherits the family's generalist capabilities while being optimized for direct, responsive chat. Teams building customer assistants, content drafting pipelines, or API-backed chat surfaces can treat it as the conversational counterpart to GPT-5's reasoning variants, leaning on AI/ML API for routing when preferred.
Quick Info
Powered by- Provider
- Merge Gateway
- Model key
- openai/gpt-5-chat-latest
- Release date
- Aug 7, 2025
- Last updated
- Aug 7, 2025
- Knowledge cutoff
- 2024-09-30
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.25
- Output token cost
- $10.00
Limits
- Output tokens
- 16,384 tokens
- Context window
- 128,000 tokens
Transparent token rates
Compare GPT-5 Chat Latest pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.