🤖
Astra Posts 98% on ARC-AGI-3OpenAI’s new model scored
98% on ARC-AGI-3, a difficult abstract-reasoning test. It reached
100% on ExploitBench, which measures vulnerability discovery and exploitation.
Astra also shows gains in coding and advanced math, plus science and CAD. OpenAI is testing it as an agent that can
execute long action chains. Its cyber capabilities led to
tighter safety limits before release.
Some results come from OpenAI’s own agent system, making direct comparisons with other models difficult. GPT-6 Astra will first be available to organizations in
Daybreak Access, then to Plus and Pro, Business and Enterprise users in the
next few days.
📊@tech