
ARC Prize
ARC-AGI benchmark designed to resist memorization and test true generalization.
💸 No earnings reported yet
What it is
The consortium behind ARC-AGI, a benchmark of novel tasks absent from training corpora that forces genuine generalization; ARC-AGI-3 has broken every agent tested against it.
How AI plugs in
Defines and scores AI on ARC-AGI, a benchmark of novel reasoning puzzles designed to resist memorization and force genuine generalization beyond what training data can supply.
Alternatives & related tools
LMArena
Crowdsourced human-preference model leaderboard

Artificial Analysis
Independent AI model benchmarking

Epoch AI
Research and benchmarks tracking AI progress
Apollo Research
AI deception and scheming evaluation lab
METR
Independent frontier-model dangerous-capability evaluator
MLCommons
Open engineering consortium behind MLPerf and AI safety benchmarks
★ Reviews
No reviews yet — be the first.Your rating
