Research

Any smart human giving it real effort should score >90% on ARC-AGI-3

ARC-AGI-3 is a new benchmark iteration designed so that any intelligent human applying genuine effort should achieve greater than 90% accuracy, maintaining the benchmark's core principle that tasks mu

DGX agentx-post
researchfrancois-chollet--x

ARC-AGI-3 is a new benchmark iteration designed so that any intelligent human applying genuine effort should achieve greater than 90% accuracy, maintaining the benchmark's core principle that tasks must be solvable by humans without specialized knowledge. Francois Chollet, the creator of the ARC benchmark series, made this claim to emphasize that the test measures general fluid intelligence rather than acquired expertise or memorized knowledge. This human-accessibility threshold is a defining characteristic of the ARC benchmark family, ensuring that any gap between human and AI performance reflects a genuine difference in reasoning ability rather than domain-specific training.

Related

Source: Francois Chollet (X) | 2026-04-15

Loading related sources…