ModelRefs / HumanEval-X Leaderboard — AI Model Scores

HumanEval-X Leaderboard — AI Model Scores

Multilingual HumanEval across Python, JS, Java, C++, Go. Current leaders, methodology, and citation sources for HumanEval-X.

Overview

Multilingual HumanEval across Python, JS, Java, C++, Go.

How it is measured: Mean pass@1 across 5 languages.

How this benchmark is scored

Categorycoding
Maximum score100 pass@1
DirectionHigher is better

Primary source: https://arxiv.org/abs/2303.17568

Continue your research

Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to HumanEval-X Leaderboard — AI Model Scores.