Open-source AI benchmarks that measure general fluid intelligence and guide AGI research.
The ARC Prize Foundation publishes the ARC-AGI benchmark series—scientifically grounded evaluations designed to reveal the gap between tasks that are easy for humans but hard for AI. The series includes ARC-AGI-3, described as the world's only unbeaten benchmark measuring agentic intelligence. The benchmarks are used by OpenAI, Anthropic, Google DeepMind, and xAI, and are recognized by NIST's Center for AI Standards and Innovation. Competitions run in partnership with Kaggle, offering over $2M in prizes for open-sourced progress on the benchmarks. ARC-AGI Benchmark Series is a product of ARC Prize Foundation.
For people
For agents
Nothing listed yet.