Currently listed through these providers:
Model details
Grok Imagine Image
Grok Imagine Image is xAI's dedicated text-to-image generation model designed to bridge the gap between natural-language descriptions and polished visual output. The model excels at producing a wide spectrum of visuals, from photorealistic scenes to stylized 3D characters and playful chibi figures, making it versatile across creative disciplines. Unlike many text-to-image systems constrained to a handful of square or portrait formats, Grok Imagine Image supports 11 preset aspect ratios—from ultra-wide cinematic banners down to TikTok verticals and Instagram posts—giving creators genuine flexibility without format compromise. The inclusion of a built-in Prompt Enhancer helps optimize raw prompts before generation, which is particularly valuable when iterating on complex visual concepts.
The model debuted as part of xAI's most comprehensive generative release to date, a multimodal stack that encompasses not just image generation but also image editing, video generation, and audio-video synchronization across five new endpoints. Grok Imagine Image operates at both 1K and 2K resolution and supports batch generation of up to 4 images per request, dramatically accelerating iteration when exploring multiple visual directions. Available through multiple API platforms, the model is positioned as a production-ready tool for developers, designers, and content teams seeking to integrate high-quality visual generation into scalable creative pipelines.
Quick Info
Powered by- Provider
- xAI
- Model key
- grok-imagine-image
- Release date
- Jan 28, 2026
- Last updated
- Jan 28, 2026
- Input modalities
- Output modalities
- Capabilities
Limits
- Output tokens
- 0 tokens
- Context window
- 8,000 tokens