On ARC-AGI-3, OpenAI's GPT-6 Astra scored 62.7 percent on ARC Prize's internal harness, surpassing the average human tester's efficiency for the first time. Epoch AI ranks it first overall with 169 points, while Artificial Analysis rates it level with predecessor Sol at 61, behind Claude Fable 5.1's 66. ARC Prize's François Chollet calls the progress "2x faster" than he expected and is moving up his AGI forecast.
No score is assigned. Sources and their independence are shown in the citation chain below.