Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Databricks logo

Model details

GPT-5.6 Luna

GPT-5.6 Luna is the economy-focused member of OpenAI's GPT-5.6 family, released alongside the flagship Sol and the balanced Terra tier. OpenAI describes Luna as the family's most cost-efficient and fastest option, and on July 30, 2026 the company cut its price by 80 percent as part of a broader push to improve price-performance across the line. DataCamp's third-party comparison frames Luna as the high-volume cost-efficient choice and notes its inheritance of programmatic tool-calling capabilities, positioning it for workloads where throughput and price matter more than top-end reasoning quality.

Through Databricks, GPT-5.6 Luna is exposed under the model key databricks-gpt-5-6-luna in the gpt-luna family on Model Serving via Foundation Model APIs pay-per-token billing, bringing the upstream OpenAI tier into the same governance and billing surface as other first-party Databricks endpoints. Databricks' July 2026 release notes confirm the July the listed price, 2026 availability, and the Foundation Model APIs limits page documents the rate-limit regime (ITPM, OTPM, QPH) that teams must size against under pay-per-token usage, with provisioned throughput as the graduation path when throttling becomes a bottleneck. For production workloads, this means Luna is best suited to cost-sensitive, high-volume pipelines where teams want OpenAI's economy tier without leaving the Databricks governance stack.

Databricksdatabricks-gpt-5-6-lunagpt-luna

Quick Info

Powered by
Provider
Databricks
Model key
databricks-gpt-5-6-luna
Release date
Jul 9, 2026
Last updated
Jul 9, 2026
Knowledge cutoff
2026-02-16
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.00
Output token cost
$6.00

Limits

Input tokens
272,000 tokens
Output tokens
128,000 tokens
Context window
400,000 tokens

Transparent token rates

Compare GPT-5.6 Luna pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GPT-5.6 Luna

Databricks

CoveragePreview

OpenAI announced a limited preview of the GPT-5.6 series on June 26, 2026, explicitly introducing Luna as "a fast and affordable model" alongside the flagship Sol and the balanced Terra tier. According to the post, Luna is positioned to bring strong capability at OpenAI's lowest cost within the family, with general ava The preview post frames GPT-5.6 Sol as the strongest model yet and notes that GPT-5.6 Luna is part of the same family rollout with strengthened safety protections for higher-risk activity, sensitive cyber requests, and repeated misuse. OpenAI indicated that an expanded suite of evaluation results and additional safety

Databricks

CoverageBenchmark

DataCamp's comparison article corroborates that GPT-5.6 Luna is the cheapest tier in OpenAI's GPT-5.6 family, which reached general availability on July 9, 2026, alongside flagship Sol and mid-tier Terra. Luna offers a 1,050,000-token context window with a 128,000-token output ceiling and inherits Programmatic Tool Cal The piece further documents that on July 30, 2026, OpenAI cut Terra's price by 20% to $2.00 input / $12 output per 1M tokens, framing Luna as the high-volume cost-efficient choice and Terra as the balanced mid-tier. This third-party coverage provides useful comparative context for developers selecting between databrick

Databricks

CoverageBenchmark

OpenAI's July 30, 2026 product post announces an 80% price reduction for GPT-5.6 Luna and a 20% reduction for GPT-5.6 Terra, effective immediately across OpenAI usage and reflected in paid Codex and ChatGPT Work subscription billing. The post positions Luna as "our fastest and most affordable model," capable of using t The same announcement introduces Fast mode in the API for GPT-5.6 Sol, delivering up to 2.5x faster speeds than Standard processing at twice the price, with backward compatibility for existing priority-tagged requests. These pricing and capability changes form the upstream context for the Databricks-hosted databricks-g

Databricks

Official sourceRelease Notes

Databricks' official July 2026 release notes (last updated July 14, 2026) confirm that OpenAI GPT-5.6 Sol, GPT-5.6 Terra, and GPT-5.6 Luna became available as Databricks-hosted models on Model Serving via Foundation Model APIs pay-per-token on July 9, 2026. The same notes also cover adjacent July platform updates, incl The release-note entry directly establishes that GPT-5.6 Luna (modelKey databricks-gpt-5-6-luna, family gpt-luna) is now reachable as a first-party Databricks-hosted endpoint, eliminating the need for customers to call OpenAI directly and bringing GPT-5.6 into the same governance and billing surface as other Foundation

Databricks

Official sourceDocumentation

The Databricks Foundation Model APIs limits and quotas page (last updated July 9, 2026) explains the rate-limit regime that governs pay-per-token endpoints such as `databricks-gpt-5-6-luna`, including input tokens per minute (ITPM), output tokens per minute (OTPM), and queries per hour (QPH), with the most restrictive For teams integrating GPT-5.6 Luna into production pipelines, this limits page is the operational reference for sizing concurrency and throughput budgets under pay-per-token billing, and for deciding when to graduate a workload to provisioned throughput to avoid 429 throttling. Pre-admission validation means clients sh

Databricks

Official sourceDocumentation

The Databricks Foundation Model APIs supported-models page (last updated July 9, 2026) is the canonical reference for GPT-5.6 Luna on Databricks, listing the endpoint name `databricks-gpt-5-6-luna` with text and image inputs, a 1.05M total token context window, and a 128K maximum output token limit. The page describes For practitioners, the supported-models entry turns the earlier release-note announcement into actionable API details: the exact endpoint string to use in requests, the multimodal input surface for vision-enabled RAG and agent workflows, and the very large 1.05M-token context that materially expands retrieval and long-

Databricks

CoverageBenchmark

Third-party benchmark aggregator llm-stats.com maintains a dedicated tracking page for GPT-5.6 Luna, listing it at rank 32 on the composite LLM Stats Score with a score of 45.1 and a blended price around $0.25 per million tokens as of mid-September 2026. The page places Luna in the Top 10% for Coding (21 of 267 tracked The llm-stats page reports a perfect 1.00/1 score on OpenAI's Connectors production benchmark for GPT-5.6 Luna, alongside tracked performance on HealthBench Consensus and other evaluations, with a Quality Tracker showing stable community voting within roughly ±2σ over the preceding 30 days. Conversation-depth data indi

Databricks

Coverage

OpenAI's July 9, 2026 general availability post for the GPT-5.6 family explicitly names GPT-5.6 Luna as the most cost-efficient tier and reports that, together with Terra, it outperforms Claude Fable 5 at around one-sixteenth the cost. The post introduces a new highest-capability "ultra" setting that coordinates multip The post includes two subsequent updates affecting Luna: on July 30, 2026, OpenAI reduced the price of GPT-5.6 Luna by 80% (and GPT-5.6 Terra by 20%), and on August 21, 2026, OpenAI dropped GPT-5.6 Sol API and credit pricing by over 20% for three months. OpenAI framed the family as delivering stronger performance per d

Videos about GPT-5.6 Luna

More models around GPT-5.6 Luna