🚀NEW LABGetting Started with Claude AgentsStart lab
Evaluation

ARC-AGI-2

First page
ARC-AGI-2
Paper summary

ARC-AGI-2 is a new benchmark designed to push the boundaries of AI reasoning beyond the original ARC-AGI. It introduces harder, more unique tasks emphasizing compositional generalization and human-like fluid intelligence, with baseline AI models performing below 5% accuracy despite strong ARC-AGI-1 results.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack