Benchmarks disagree on GPT-6 Astra, but its human-beating efficiency on ARC-AGI-3 pulls Chollet’s AGI forecast forward
The Decoderen

OpenAI's GPT-6 Astra is drawing contradictory benchmark verdicts. Epoch AI puts it out in front with 169 points, while Artificial Analysis rates it no better than its predecessor and behind Claude Fable 5.1. The biggest surprise comes from ARC-AGI-3, where Astra works more efficiently than the average human for the first time. ARC Prize chief François Chollet doesn't call this proof of AGI, but he does see the progress running "twice as fast" as he expected, and he's moving up his AGI forecast. The article Benchmarks disagree on GPT-6 Astra, but its human-beating efficiency on ARC-AGI-3 pulls Chollet’s AGI forecast forward appeared first on The Decoder .
This is a short summary published by AI Global Wire. The full article is owned and hosted by The Decoder — open it there to read it in full.
Read the full story at The Decoder- OpenAI
- Anthropic
- Verktyg
Related AI news
- A20 Pro fürs iPhone: Gerüchte um GPU-Anzahl, Speicherbandbreite und Cacheheise online – KI · September 4, 2026
- Borde AI vara en valfråga?Computer Sweden · September 4, 2026
- Instagram’s AI detection is a mess (again)The Verge AI · September 4, 2026
- [Ekstra] Sikkerhetsselskap slår alarm om autonome KI-angrepdigi.no · September 4, 2026
- Applen ex-työntekijän koneelta löytyi jotain odottamatonta – Vakoilusyytökset kiihtyvätTivi · September 4, 2026
- Why AI food looks like thatThe Verge AI · September 4, 2026