Measuring Human Performance on ARC-AGI-3 AGI is here when a system can learn like a human. However there is still a gap between what humans can learn and what AI can learn. ARC Prize Foundation exists...
Announcing ARC-AGI-3 A New Challenge for Frontier Agentic Intelligence Today we're excited to announce the release of ARC-AGI-3, a series of hundreds of interactive environments and thousands of game-...
ARC Prize 2025 Results & Analysis Year of the Refinement Loop We've officially wrapped Year 2 of ARC Prize! While the Grand Prize remains unclaimed, we're excited to announce the ARC Prize 2025 Score ...
Announcing ARC Prize Verified Today we're announcing **ARC Prize Verified**, a program to increase the rigor of evaluating frontier systems on the ARC-AGI benchmark. In addition to certified score ver...
The Hidden Drivers of HRM's Performance on ARC-AGI We scored on hidden tasks, ran ablations, and found that performance comes from an unexpected source On June 8, 2025, the Hierarchical Reasoning Mode...
ARC-AGI-3 Preview: 30-day learnings Highlighting the gap between humans and AI with Interactive Benchmarks ARC-AGI-3, the first Interactive Reasoning Benchmark by ARC Prize Foundation On July 17, we r...
ARC Prize Foundation Statement on the US AI Action Plan Last week, the White House released its AI Action Plan. We’re encouraged to see the White House take transparency and measurement seriously as c...
We tested every major AI reasoning system. There is no clear winner. Reflecting on Frontier AI Reasoning Systems It's been six months since OpenAI achieved a breakthrough high score on ARC-AGI-1 with....
Analyzing o3 and o4-mini with ARC-AGI ARC Prize Foundation is a nonprofit committed to serving as the **North Star for AGI** by building open reasoning benchmarks that highlight the gap between what’s...