2025
M-IFEval: Multilingual Instruction-Following Evaluation
NAACL 2025findings
Instruction following is a core capability of modern Large language models (LLMs), making evaluating this capability essential to understanding these models. The Instruction Following Evaluation (IFEval) benchmark from the literature does this using objective criteria, offering a measure of LLM perf…