LLM Releases
← Catalog

GPT-6 Astra

Available
OpenAIFrontierProprietary

OpenAI's new frontier flagship, released Sep 3, 2026 (API id gpt-6-astra) as the successor to GPT-5.6 Sol and OpenAI's most capable and most aligned model to date. Positioned around three areas: state-of-the-art computer/browser use, a step change in producing finished professional artifacts (documents, slides, spreadsheets, and websites via Sites in ChatGPT that follow the user's templates), and a jump in cybersecurity capability that crosses the 'Critical' threshold of OpenAI's Preparedness Framework. It decides when to ask clarifying questions versus proceed on assumptions, holds constraints across mid-task steering, and — with an updated Codex harness — keeps searchable notes across context windows instead of lossy compaction. Carries a 1M-token context window; available as gpt-6-astra in the OpenAI API and on Amazon Bedrock (with Microsoft Azure), plus a GPT-6 Astra Pro tier for ChatGPT Pro, Business, and Enterprise (off by default at launch, enabled per workspace). Standard API pricing is $10/$50 per Mtok input/output, with a Fast mode at ~2x price (~$20/$100) for up to 2.5x speed, plus separate cache rates. Vendor-reported benchmarks include OSWorld 2.0 72.6% (at ~47% less time per task than Sol), FrontierMath Tier 4 v2 97.6%, GPQA Diamond 96.0%, Terminal-Bench 4.0 57.7%, and ExploitBench 100%; the marquee ARC-AGI-3 ~99.9% depends on a stateful provider-adapter harness (stateless runs score far lower), and it trails Claude Fable 5.1 on Humanity's Last Exam with tools (57.2% vs 65.0%). Because it crosses the Critical cyber threshold, exploit-creation capabilities ship gated behind OpenAI's Daybreak program; OpenAI also flags a regression in chain-of-thought monitorability as a research priority.

Specifications

License
Proprietary
Weights
Not released
Architecture
unknown
Parameters
Undisclosed
Context window
1M tokens
Max output
128K tokens
Knowledge cutoff
Price (in / out, $/M)
$10 / $50
Modalities
TextVisionCode

Benchmarks

No benchmark scores recorded yet. Spotted some? Submit a correction.

Vendor-reported figures are claims until independently verified. See methodology.