ModelRefs / LiveCodeBench Leaderboard — AI Model Scores

LiveCodeBench Leaderboard — AI Model Scores

Fresh competitive programming problems with timestamped contamination guard. Current leaders, methodology, and citation sources for LiveCodeBench.

Overview

Fresh competitive programming problems with timestamped contamination guard.

How it is measured: pass@1 on problems released after model training cutoff.

How this benchmark is scored

Categorycoding
Maximum score100 pass@1
DirectionHigher is better
Evidence depthcomplete

Primary source: https://livecodebench.github.io/

Published results

ModelScore
GPT-572
DeepSeek R165.9

Each score reflects the protocol and date of its own source run. Results from different harnesses are not directly comparable.

Continue your research

Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to LiveCodeBench Leaderboard — AI Model Scores.