Announcing ARC-AGI-3 The only unsaturated agentic intelligence...

The only unsaturated agentic intelligence benchmark in the world
Humans score 100%, AI <1%
This human-AI gap demonstrates we do not yet have AGI
Most benchmarks test what models already know, ARC-AGI-3 tests how they learn
No instructions, Core Knowledge Priors-only
In order to beat these games, AI must:
• Explore the environment
• Form hypotheses
• Execute a plan
• Learn and adapt
Key failure modes seen in our early testing:
• Thinking it is playing another game
• Holding on to early hypothesis
• Unable to forecast into the future
Both AI + human runs have sharable replays
Watch Gemini 3.1 do well on some games, poorly on others:
arcprize.org/replay/34a9614…
arcprize.org/replay/d0e0768…
Get involved:
• Play a Game: arcprize.org/tasks/ls20
• Build Agents: docs.arcprize.org
• Win Prizes: arcprize.org/competitions/2…


