ModelRefs / Open LLM Leaderboard v2 Leaderboard — AI Model Scores

Open LLM Leaderboard v2 Leaderboard — AI Model Scores

HuggingFace composite ranking of open-weight models across 6 normalized sub-benchmarks. Current leaders, methodology, and citation sources for Open LLM Leaderboard v2.

Overview

HuggingFace composite ranking of open-weight models across 6 normalized sub-benchmarks.

How it is measured: Average of IFEval + BBH + MATH Lvl5 + GPQA + MuSR + MMLU-Pro; all zero-shot or few-shot per benchmark spec.

How this benchmark is scored

Categoryopen-source
Maximum score100 avg score
DirectionHigher is better

Primary source: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard

Continue your research

Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to Open LLM Leaderboard v2 Leaderboard — AI Model Scores.