Sulat.com
AI models
DevPass (LLM Gateway) logo

Model details

GPT-6 Astra

GPT-6 Astra is positioned as OpenAI's flagship model for demanding end-to-end work, with a design emphasis on long-horizon agentic behavior that involves computer and browser interaction rather than only single-turn chat. The model fits workflows such as advanced analysis, software engineering, deep research, scientific work, and document creation, where sustained tool use and multi-step planning matter. OpenRouter's listing frames it explicitly around these use cases, and a Lenny's Newsletter hands-on report from early-access testing highlights noticeable gains on standing tasks that prior models could not solve in one shot. That combination of agentic framing and independent user feedback points to a model intended to operate as a capable collaborator on extended projects rather than a narrow specialist.

In practical terms, GPT-6 Astra brings a very large context window suitable for ingesting substantial codebases, long documents, or research corpora in a single session, which pairs naturally with its computer-use and browser-use strengths. The model is offered across multiple hosting routes on OpenRouter, including a Flex tier at lower cost and standard tiers on Azure and OpenAI infrastructure, giving teams flexibility to trade off price, latency, and throughput. Early commentary points to large jumps in mathematical ability and broad knowledge work, with the model described as fast and less prone to the irritating failure modes of its predecessors. For teams building agent pipelines, coding copilots, or research assistants that need sustained tool orchestration, GPT-6 Astra presents itself as a general-purpose flagship with a clear bias toward longer, more autonomous tasks.

DevPass (LLM Gateway)gpt-6-astragpt

Quick Info

Powered by
Provider
DevPass (LLM Gateway)
Model key
gpt-6-astra
Release date
Sep 3, 2026
Last updated
Sep 3, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$10.00
Output token cost
$50.00

Limits

Output tokens
1,050,000 tokens
Context window
1,050,000 tokens

Transparent token rates

Compare GPT-6 Astra pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GPT-6 Astra

DevPass (LLM Gateway)

CoverageBenchmark

ComputingForGeeks' hands-on review confirms OpenAI shipped GPT-6 Astra on September 3, 2026 with day-two API access via the OpenAI-compatible endpoint serving the 'gpt-6-astra' model. The piece documents detailed specs: a 1,050,000-token context window (922K input / 128K output), a knowledge cutoff of April 30, 2026 pe The article documents GPT-6 Astra's pricing structure: $10 per million standard input tokens, $1 per million cached input reads, $12.50 per million cache writes, and $50 per million output tokens, with batch pricing at $5/$25 per million input/output. Long-context usage above 272K tokens steps up to $20/$75 per million

DevPass (LLM Gateway)

Coverage

Kingy.ai's rolling-updates tracker documents the operational rollout mechanics for GPT-6 Astra across multiple surfaces. Microsoft Foundry now lists 'gpt-6-astra' version 2026-09-03 in its model catalog with a live region matrix, where Tier 5 and Tier 6 Azure subscriptions receive quota by default while lower tiers mus The piece records token-based credit rates for ChatGPT Work and Codex usage of GPT-6 Astra: 250 credits per million input tokens, 25 per million cached input tokens, and 1,250 per million output tokens, with Fast mode costing 2.5x the Standard credit rate in Work and Codex versus 2x in the OpenAI API. Pro $100, Pro $20

DevPass (LLM Gateway)

CoverageRelease Notes

OpenAI's official release notes dated September 3, 2026 announce GPT-6 Astra's general availability with the model ID 'gpt-6-astra', exposed on both v1/responses and v1/chat/completions endpoints. Access is rolling out initially to a limited set of organizations, with broader availability planned over the coming days. The release notes enumerate concrete API-breaking changes for developers migrating to GPT-6 Astra: the model no longer supports the 'none' reasoning effort level, custom temperature or top_p values, or log probabilities; tool calling is now Responses-API-only. The Responses API gains three new controls relevant to long

Videos about GPT-6 Astra

More models around GPT-6 Astra