ModelRefs / Humanity's Last Exam Leaderboard — AI Model Scores

Humanity's Last Exam Leaderboard — AI Model Scores

Multi-disciplinary expert-level exam, contamination resistant. Current leaders, methodology, and citation sources for Humanity's Last Exam.

Overview

Multi-disciplinary expert-level exam, contamination resistant.

How it is measured: Closed-book; >3k questions across 100+ domains.

How this benchmark is scored

Categoryreasoning
Maximum score100 % accuracy
DirectionHigher is better

Primary source: https://lastexam.ai/

Continue your research

Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to Humanity's Last Exam Leaderboard — AI Model Scores.