← Browse

World's Fair 2025

Measuring AGI: Interactive Reasoning Benchmarks for ARC-AGI-3 — Greg Kamradt, ARC Prize Foundation

Greg Kamradt , ARC Prize Foundation

Overview

This talk introduces ARC-AGI-3, a new benchmark designed to measure artificial general intelligence (AGI) by focusing on interactive reasoning and skill acquisition efficiency. The benchmark aims to create problems that are solvable by humans but challenging for current AI, thereby guiding AI research and development towards human-level intelligence. It moves beyond single-turn, static benchmarks to simulate more realistic, open-world exploration and learning scenarios.

Who should watch

Key takeaways

Notable quotes

*Intelligence is skill acquisition efficiency.*
*As long as we can come up with problems that humans can still do but machines cannot, I would again assert that we do not have AGI.*

Watch on YouTube →

Unofficial community note. Prefer the recording for nuance.