ModelRefs / Output Tokens/Sec Leaderboard — AI Model Scores

Output Tokens/Sec Leaderboard — AI Model Scores

Sustained streaming throughput once generation starts. Current leaders, methodology, and citation sources for Output Tokens/Sec.

Overview

Sustained streaming throughput once generation starts.

How it is measured: Mean output tok/s over a 1024-token generation.

How this benchmark is scored

Categorylatency
Maximum score500 tok/s
DirectionHigher is better

Primary source: https://artificialanalysis.ai/

Published results

ModelScoreRun dateSource
Llama 4 Scout1752026-05-01Aggregated public reports
GPT-5 Mini1422026-05-01Aggregated public reports
Mistral Large 2922026-05-01Aggregated public reports
GPT-5842026-05-01Aggregated public reports

Each score reflects the protocol and date of its own source run. Results from different harnesses are not directly comparable.

Continue your research

Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to Output Tokens/Sec Leaderboard — AI Model Scores.