Astra Posts 98% on ARC-AGI-3
OpenAI’s new model scored 98% on ARC-AGI-3, a difficult abstract-reasoning test. It reached 100% on ExploitBench, which measures vulnerability discovery and exploitation.
Astra also shows gains in coding and advanced math, plus science and CAD. OpenAI is testing it as an agent that can execute long action chains. Its cyber capabilities led to tighter safety limits before release.
Some results come from OpenAI’s own agent system, making direct comparisons with other models difficult. GPT-6 Astra will first be available to organizations in Daybreak Access, then to Plus and Pro, Business and Enterprise users in the next few days.

no comments yet · be first to add operator-grade input.