📖 The AI Tool Bible
reasoning

ARC-AGI

Francois Chollet's Abstraction and Reasoning Corpus. Grid-based visual puzzles designed to be trivial for humans but hard for machines that memorised the internet. The most publicised general-intelligence benchmark of 2024-25.

Official leaderboard →

Top scores

#ModelScore
1o387.5% (compute-heavy)
2gpt-532.0%
3claude-opus-4-828.5%
4gemini-2-5-pro15.0%

Scores are snapshots from public leaderboards at the time of last update. Follow the source link for the live board.