ModelRefs / DS-1000 Leaderboard — AI Model Scores

DS-1000 Leaderboard — AI Model Scores

Data-science code synthesis across NumPy, Pandas, PyTorch, scikit-learn. Current leaders, methodology, and citation sources for DS-1000.

Overview

Data-science code synthesis across NumPy, Pandas, PyTorch, scikit-learn.

How it is measured: Pass@1 with test-driven evaluation.

How this benchmark is scored

Categorycoding
Maximum score100 pass@1
DirectionHigher is better

Primary source: https://ds1000-code-gen.github.io/

Continue your research

Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to DS-1000 Leaderboard — AI Model Scores.