← Back to Model Beat
Models·Sep 3·all news from September 3, 2026

Benchmarks disagree on GPT-6 Astra, but its human-beating efficiency on ARC-AGI-3 pulls Chollet’s AGI forecast forward

Artificial Analysis updated its Intelligence Index to version 4.2 following criticism regarding how it evaluated the performance of GPT-6 Astra. The revised scoring places GPT-6 Astra four points ahead of its previous iteration, though the model remains behind Claude Fable 5.1. This adjustment addresses concerns over the accuracy of the platform's benchmarking methodology for evaluating modern AI capabilities.

ModelsGPT-6 Astra

Covered by 3 sources · 7 articles

Related stories

ModelsGPT-6 Astra: The next generation in intelligence for workSep 7 · 8 sourcesModelsPerplexity trusts GPT-6 Astra with end-to-end systemsSep 12 · 4 sourcesModelsSakana AI has released 'Fugu Ultra v2,' which surpasses the GPT-6 Astra by utilizing multiple models, and has also introduced the cost-effective 'Fugu Max.'Sep 14 · 3 sourcesOpinionCognition helps Devin test its own work with GPT‑6 AstraSep 11