ModelRefs / Output Tokens/Sec Leaderboard — AI Model Scores

Output Tokens/Sec Leaderboard — AI Model Scores

Sustained streaming throughput once generation starts. Current leaders, methodology, and citation sources for Output Tokens/Sec.

Overview

Sustained streaming throughput once generation starts.

How it is measured: Mean output tok/s over a 1024-token generation.

How this benchmark is scored

Categorylatency
Maximum score500 tok/s
DirectionHigher is better
Evidence depthincomplete

Primary source: https://artificialanalysis.ai/

Published results

ModelScore
Llama 4 Scout175
GPT-5 Mini142
Mistral Large 292
GPT-584

Each score reflects the protocol and date of its own source run. Results from different harnesses are not directly comparable.

Continue your research

Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to Output Tokens/Sec Leaderboard — AI Model Scores.