Currently listed through these providers:
Model details
OpenAI: GPT-5 Image ($$$$)
GPT-5 Image is a natively multimodal model that merges OpenAI's GPT-5 Mini language architecture with GPT Image 1 Mini's image generation capabilities. This combination enables the model to ingest file, image, and text inputs while producing both image and text outputs. Built around a 400,000-token context window, the architecture is engineered to handle complex visual and language interactions, excelling particularly at instruction following, text rendering inside images, and detailed image editing. The design intent centers on unified multimodal reasoning where visual understanding and language generation reinforce each other.
Training and post-training development emphasize alignment between language comprehension and visual generation, contributing to the model's strong instruction-following performance. On the LMArena Text-to-Image Human Preference benchmark, GPT-5 Image achieved an Elo score of 1,264, placing it ahead of comparable models in blind comparisons. This measured preference reflects the practical payoff of combining GPT-5 language capabilities with GPT Image 1's visual synthesis, making the model especially suited for workflows that demand coherent visual creation alongside nuanced text understanding.
Quick Info
Powered by- Provider
- Kilo Gateway
- Model key
- openai/gpt-5-image
- Release date
- Oct 14, 2025
- Last updated
- Oct 14, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $10.00
- Output token cost
- $10.00
Limits
- Output tokens
- 128,000 tokens
- Context window
- 400,000 tokens
Transparent token rates
Compare OpenAI: GPT-5 Image ($$$$) pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about OpenAI: GPT-5 Image ($$$$)
No articles yet. Fetch the latest news to show it here.