2026
LexInstructEval: Lexical Instruction Following Evaluation for Large Language Models
AAAI 2026technical
The ability of Large Language Models (LLMs) to precisely follow complex and fine-grained lexical instructions is a cornerstone of their utility and controllability. However, evaluating this capability remains a significant challenge. Current methods either rely on subjective and costly human evaluat