Currently listed through these providers:
Model details
GPT-5.6 Luna Pro
GPT-5.6 Luna Pro is positioned as a reasoning-focused variant of the base GPT-5.6 Luna model, delivered through OpenRouter with its reasoning mode set to "pro." According to the OpenRouter listing, the model shares the same underlying weights as the standard Luna version, but routes requests through a higher-effort reasoning configuration intended to yield stronger answers on complex tasks, with a link to OpenAI's published reasoning-mode guide for implementation details. This makes the Pro tier best understood not as a separate architecture but as a serving profile that trades latency for depth of deliberation, suiting workloads where answer quality on hard prompts matters more than raw speed.
The model is built to handle very large contexts, with OpenRouter advertising an approximately one-million-token window and the third-party Bifrost listing corroborating a 1,050K-token input ceiling paired with the cataloged API limit maximum output. It accepts text and image inputs and emits text, and supports the practical tooling that long-context applications need: tool calling, structured response schemas, and prompt caching, all confirmed by Bifrost's comparison metadata. Together, the wide that quick-info value, multimodal input, and first-class function-calling and caching make the Pro variant a natural fit for document-heavy agents, multi-turn analytical workflows, and any pipeline that needs to hold large codebases, transcripts, or mixed media in a single conversation while orchestrating external tools.
Quick Info
Powered by- Provider
- OpenRouter
- Model key
- openai/gpt-5.6-luna-pro
- Release date
- Jul 9, 2026
- Last updated
- Jul 9, 2026
- Knowledge cutoff
- 2026-02-16
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.20
- Output token cost
- $1.20
Limits
- Input tokens
- 922,000 tokens
- Output tokens
- 128,000 tokens
- Context window
- 1,050,000 tokens