GPT-6 Astra scores 62.7% on ARC-AGI-3 (99.9% with provider harness), outacting humans on 96% of levels
ARC Prize reported on 2026-09-03 that OpenAI's GPT-6 Astra scored 62.7% on ARC-AGI-3 (semi-private) with the standard harness ($26K) and 99.9% ($19K) with OpenAI's own provider-adapter harness, using fewer actions than the human baseline on 96% of levels; ARC Prize will now label both conditions separately.
Key facts
- Standard harness: 62.7% at $26,098; Provider Adapter harness: 99.9% at $18,817
- Fewer actions than human baseline on 96.0% of levels; 51.7% fewer actions per level on average (provider harness)
- Human participants were paid ~ $12.78 per attempted game
- Other ARC-AGI-3 scores: Claude Opus 5 30.16% (Jul 24), Gemini 3.8 Flash 35.00%, GPT-5.6 7.78%, Grok 4.6 2.11% (leaderboard as of late Sept)
- Same leaderboard: GPT-6 95.0% on ARC-AGI-2; Claude Opus 5.5 93.3% (Sep 22)
- ARC Prize is exploring next-generation benchmarks (recursive self-improvement, open-ended innovation)
What happened
Six months after ARC-AGI-3 launched with frontier models near 0%, GPT-6 Astra reached 62.7% under the neutral harness. With OpenAI's context-management setup it reached 99.9%, a result the shared harness did not reproduce, so ARC Prize now reports both. ARC Prize said Astra "builds the most precise symbolic model of novel environments we've seen."
Why it matters
ARC-AGI-3 was meant to measure human-like skill acquisition; its near-saturation (and the harness gap) shows both how fast agentic reasoning improved in 2026 and how much scaffolding now drives scores.
Changelog
- 2026-09-29: added post link(s) (OpenAI cluster post research)
- 2026-09-29: created
Related posts (4)
- Chollet: GPT-6 Astra is a 'step-function change' on ARC-AGI-3 François Chollet @fchollet · x · 2026-09-03
The ARC-AGI creator confirms near-saturation of ARC-AGI-3 roughly twice as fast as he predicted, while declining to call it AGI. - OpenAI: 'This is GPT-6 Astra' OpenAI @OpenAI · x · 2026-09-03
OpenAI's official launch post for GPT-6 Astra, the model OpenAI leadership framed as the start of the AGI era. - Hot take on OpenAI GPT-6 Astra, with a challenge to Brockman's AGI claims Gary Marcus @GaryMarcus · x · 2026-09-03
The leading LLM skeptic called Astra a genuine advance and a vindication of symbolic world models, while rejecting Greg Brockman's claim that it is AGI. - ARC Prize ARC Prize @arcprize · x · 2026-09-03
Cited as a source by: 2026-09-03-arc-agi-3-gpt-6-astra
Related events
- ARC Prize launches ARC-AGI-3, an interactive game benchmark where frontier AI scored under 1% ★★★★
- Jensen Huang declares "AGI has arrived" with GPT-6 Astra; Greg Brockman: "we're now moving into the AGI era" ★★★★
Sources (5)
- officialARC Prize: OpenAI's GPT-6 Astra on ARC-AGI-3
- officialARC Prize results leaderboard
- discussionARC Prize on X
- press36Kr: GPT-6 scores 99.9%, ARC exam forced remake
- discussionFrançois Chollet on X: Astra a 'step-function change' on ARC-AGI-3
id: 2026-09-03-arc-agi-3-gpt-6-astra · updated 2026-09-29 · open in the interactive timeline