Currently listed through these providers:
Model details
GPT-5.6 Luna (EU)
GPT-5.6 Luna sits inside the broader gpt-luna family as a budget-oriented sibling designed for organizations that need quick, repeatable inference at scale rather than maximum depth of reasoning. Catalog positioning describes it as a cost-efficient preview variant optimized for high-volume workloads, which aligns with its placement between lighter throughput options and the flagship agentic coding models in the same generation. The model retains multimodal input handling while restricting generation to text, making it well suited to document and image-driven pipelines that produce structured outputs through tool calls and JSON-formatted responses.
Practically, GPT-5.6 Luna fits scenarios where consistent throughput, predictable token costs, and the ability to attach files or screenshots matter more than frontier-level problem solving. Teams running batch summarization, customer support triage, routine code assistance, or content transformation can route requests through EU-hosted endpoints to keep latency and data residency predictable for European users. Its support for tool calling and structured output lets it plug into orchestration layers that automate follow-on actions, while the closed weights keep deployment to API access only, leaving providers to handle scaling and compliance concerns.
Quick Info
Powered by- Provider
- Requesty
- Model key
- gpt-5.6-luna@eu
- Release date
- Jul 9, 2026
- Last updated
- Jul 9, 2026
- Knowledge cutoff
- 2026-02-16
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.22
- Output token cost
- $1.32
Limits
- Input tokens
- 922,000 tokens
- Output tokens
- 128,000 tokens
- Context window
- 1,050,000 tokens
Latest news about GPT-5.6 Luna (EU)
Videos about GPT-5.6 Luna (EU)
Recent tweets and retweets from Requesty
More models around GPT-5.6 Luna (EU)
This exact model name is also listed by 35 other providers.