Sulat.com
AI models
Kenari logo

Model details

GPT-Image-2

GPT-Image-2 is positioned as a generation model that blends diffusion-style image synthesis with reasoning-driven planning, integrating OpenAI's O-series "Thinking mode" to lay out compositions, search the web, and pull from uploaded documents before rendering a final frame. An "Instant" path handles everyday prompts while a deeper Thinking tier, including a Pro-only ImageGen Pro layer, is aimed at more deliberate editorial work. The intended use case is asset-grade creative production: posters, infographics, editorial spreads, and localized marketing where typography and layout integrity matter as much as the picture itself.

In practical terms, the model is built around three differentiators. First, text rendering inside images is treated as a headline capability rather than a side effect, with cleaner typography for headings, captions, and infographic labels. Second, multilingual coverage is strengthened across Japanese, Korean, Chinese, Hindi, and Bengali, which makes it a better fit for agencies producing region-specific creative. Third, the editing loop is conversational and flexible, supporting selective area edits, any aspect ratio from 3:1 to 1:3, up to eight consistent multi-image outputs with character continuity, and a maximum 4K output resolution, giving creative teams room to iterate from rough concept to deliverable without leaving the same workflow.

Kenarigpt-image-2gpt-image

Quick Info

Powered by
Provider
Kenari
Model key
gpt-image-2
Release date
Apr 21, 2026
Last updated
Apr 21, 2026
Input modalities
Output modalities
Capabilities

Cost

A provider subscription or plan supersedes token-based pricing for this model.

Limits

Output tokens
16,384 tokens
Context window
272,000 tokens

Latest news about GPT-Image-2

Videos about GPT-Image-2

More models around GPT-Image-2