GPT-6 Astra AGI Claims Face Benchmark Skepticism
OpenAI's GPT-6 Astra initially claimed to have achieved AGI by replicating games, but subsequent analysis suggests its high benchmark scores likely stem from harness manipulation rather than true general intelligence. The saga currently stands with skepticism mounting that the model's performance reflects test exploitation instead of genuine cognitive breakthroughs.
-
OpenAI launches Astra for Law, a GPT-6 Astra variant with legal search and Trusted Access
OpenAI announced Astra for Law, a specialized configuration of its GPT-6 Astra model that couples a legal search index of more than 230 million URLs with workflow tools and strict access controls…
1 source -
OpenAI’s GPT‑6 Astra 99.9% ARC‑AGI‑3 score may reflect benchmark harness, not AGI
OpenAI announced that its GPT‑6 Astra achieved a 99.9% score on the ARC‑AGI‑3 benchmark, a figure that quickly became a headline for the company’s push toward an AGI narrative. The nonprofit ARC…
1 source -
OpenAI’s GPT‑6 Astra Claims AGI, Quickly Replicates Existing Games
OpenAI has unveiled GPT‑6 Astra, a large‑language model that the company says is its most intelligent and aligned system yet. The announcement comes with a series of demos in which Astra is shown…
2 sources