News
AI Summary
4 Sept 202623 Rabiʻ I 1448 AH
Benchmarks disagree on GPT-6 Astra, but its human-beating efficiency on ARC-AGI-3 pulls Chollet’s AGI forecast forward

Benchmarks disagree on GPT-6 Astra, but its human-beating efficiency on ARC-AGI-3 pulls Chollet’s AGI forecast forward

OpenAI's GPT-6 Astra is experiencing mixed benchmark evaluations, scoring 169 points from Epoch AI, while Artificial Analysis ranks it below its predecessors. The most surprising result comes from the ARC-AGI-3 test, where Astra outperformed the average human for the first time. ARC Prize chief François Chollet does not view this as proof of AGI, but he acknowledges the model's progress is running 'twice as fast' as he anticipated.

Follow these topics

Sign in to follow the topics that matter to you

Sign in to follow

This summary is generated with AI and receives periodic editorial review. Refer to the original source for full details.

0
0 reading now

Insight Score

Rate to unlock

Sign in to react, rate, and save. Sign In