ModelRefs / IFEval Leaderboard — AI Model Scores

IFEval Leaderboard — AI Model Scores

Instruction-following evaluation with verifiable constraints. Current leaders, methodology, and citation sources for IFEval.

Overview

Instruction-following evaluation with verifiable constraints.

How it is measured: Strict prompt-level and instruction-level accuracy.

How this benchmark is scored

Categoryreasoning
Maximum score100 % accuracy
DirectionHigher is better

Primary source: https://arxiv.org/abs/2311.07911

Continue your research

Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to IFEval Leaderboard — AI Model Scores.