ModelRefs / Humanity's Last Exam Leaderboard — AI Model Scores
Humanity's Last Exam Leaderboard — AI Model Scores
Multi-disciplinary expert-level exam, contamination resistant. Current leaders, methodology, and citation sources for Humanity's Last Exam.
Overview
Multi-disciplinary expert-level exam, contamination resistant.
How it is measured: Closed-book; >3k questions across 100+ domains.
How this benchmark is scored
| Category | reasoning |
|---|---|
| Maximum score | 100 % accuracy |
| Direction | Higher is better |
Primary source: https://lastexam.ai/
Continue your research
Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to Humanity's Last Exam Leaderboard — AI Model Scores.