OPENAI ๐Ÿ”ฅ: Astra scored 63% on ARC-AGI-3 with a standard harness.

๐Ÿšจ AI News | TestingCatalog

๐Ÿšจ AI News | TestingCatalog

@testingcatalog

Latest AI News on AI Agents, Model Releases, Tools, Leaks, and Rumors ๐Ÿ—ž

7,502 ืžื ื•ื™ื™ื
ืคืชื— ื‘ื˜ืœื’ืจื
OPENAI ๐Ÿ”ฅ: Astra scored 63% on ARC-AGI-3 with a standard harness.

> GPT-6 Astra surpasses the human baseline in action efficiency on ARC-AGI-3. It used fewer actions than the median tested human on 96% of levels.

> A key behavior observed in GPT-6 Astra was its ability to turn unfamiliar environments into compact symbolic world models. It represented game mechanics as logical rules and developed its own domain-specific language shorthand to track state and plan actions.
ืคืชื— ืืช ื”ืคื•ืกื˜ ื‘ื˜ืœื’ืจื