Early preview
O

OpenAI

GPT-6 Astra

Announced Sep 3, 2026

A new generation for complex work

01

What could it do?

OpenAI announced improved computer use, software engineering, science and professional reasoning.

02

What changed?

Combined advances in pretraining, reinforcement learning and alignment.

WHY IT MATTERED

Extended the range of difficult tasks supported by a general model.

03

Where it fell short

Rollout was phased. OpenAI’s internal math research model must not be assumed to be this public model.

Sources checked Oct 7, 2026

Release facts were checked against the sources below. Performance claims belong to the developers; we have not independently tested these models.

Developer announcement or release log

Benchmark results

EPOCH AI CAPABILITY ESTIMATE
166.4index points
Tested variant: GPT-6 AstraSource interval: 163.3 to 171.5Variant date in source: Sep 3, 2026

A benchmark estimate, not a percentage or capability multiplier. Reasoning settings are not specified in this source table. Historical estimates can change in later snapshots.

Epoch AI methodology ↗Download the source snapshotChecked Oct 7, 2026 · CC BY 4.0
PUBLISHED BENCHMARK RESULT
53.0index points
Tested: GPT-6 Astra (max)Intelligence Index v4.3.2

Selected highest effort variant where available in the current table. Starred partial results are excluded. Claude fallback configurations can use other models when safeguards intervene.

Source: Artificial Analysis Intelligence Index ↗Download selected resultsChecked Oct 8, 2026
PUBLISHED BENCHMARK RESULT
1475rating points
Tested: gpt-6-astra-maxText Arena Overall · October 8, 2026Reported interval: 1468 to 148210,536 votes

One explicitly named variant per release. Scores come from the same Overall snapshot; preliminary entries and reported intervals are preserved. These are current ratings of earlier variants, not their launch day ratings.

Source: Text Arena ↗Download selected resultsChecked Oct 8, 2026
FOLLOW WHAT HAPPENS NEXT

Breakthroughs, with the followup.

A weekly brief on new discoveries, meaningful checks and what you can actually use.