ModelRefs / LiveCodeBench Leaderboard — AI Model Scores
LiveCodeBench Leaderboard — AI Model Scores
Fresh competitive programming problems with timestamped contamination guard. Current leaders, methodology, and citation sources for LiveCodeBench.
Overview
Fresh competitive programming problems with timestamped contamination guard.
How it is measured: pass@1 on problems released after model training cutoff.
How this benchmark is scored
| Category | coding |
|---|---|
| Maximum score | 100 pass@1 |
| Direction | Higher is better |
| Evidence depth | complete |
Primary source: https://livecodebench.github.io/
Published results
| Model | Score |
|---|---|
| GPT-5 | 72 |
| DeepSeek R1 | 65.9 |
Each score reflects the protocol and date of its own source run. Results from different harnesses are not directly comparable.
Continue your research
Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to LiveCodeBench Leaderboard — AI Model Scores.