Currently listed through these providers:
Model details
GPT-5.6 Luna
GPT-5.6 Luna is presented through the Microsoft Foundry catalog as a frontier-class model aimed at complex professional workloads that demand both responsiveness and depth. The listing frames it as the most capable offering in its series, built to deliver faster and more reliable results than previous generations, and to scale across multilingual and large-context applications without sacrificing throughput. That positioning suggests an emphasis on low-latency inference and dependable reasoning quality, making the model a fit for production pipelines where response time, consistency, and broad coverage matter as much as raw capability.
Within the Foundry portfolio, GPT-5.6 Luna is offered as a Direct from Azure model, meaning Microsoft manages hosting, billing, and governance rather than routing requests to a third party, which simplifies procurement and supports enterprise compliance requirements. The catalog copy emphasizes cost efficiency at scale and scalable reasoning that improves with available compute, pointing to a design that rewards larger that quick-info value and longer thinking budgets while keeping per-token economics attractive for high-volume use. For teams building agentic systems, document analysis workflows, or interactive assistants over very long inputs, this combination of managed delivery and frontier-grade reasoning is the practical appeal of the model.
Quick Info
Powered by- Provider
- Neon
- Model key
- gpt-5-6-luna
- Release date
- Jul 9, 2026
- Last updated
- Jul 9, 2026
- Knowledge cutoff
- 2026-02-16
- AI SDK package
@ai-sdk/openai- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.00
- Output token cost
- $6.00
Limits
- Input tokens
- 922,000 tokens
- Output tokens
- 128,000 tokens
- Context window
- 1,050,000 tokens
Latest news about GPT-5.6 Luna
Videos about GPT-5.6 Luna
More models around GPT-5.6 Luna
This exact model name is also listed by 35 other providers.