ModelRefs / HumanEval-X Leaderboard — AI Model Scores
HumanEval-X Leaderboard — AI Model Scores
Multilingual HumanEval across Python, JS, Java, C++, Go. Current leaders, methodology, and citation sources for HumanEval-X.
Overview
Multilingual HumanEval across Python, JS, Java, C++, Go.
How it is measured: Mean pass@1 across 5 languages.
How this benchmark is scored
| Category | coding |
|---|---|
| Maximum score | 100 pass@1 |
| Direction | Higher is better |
Primary source: https://arxiv.org/abs/2303.17568
Continue your research
Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to HumanEval-X Leaderboard — AI Model Scores.