AAAI 2026technical0 citations
LexInstructEval: Lexical Instruction Following Evaluation for Large Language Models
Huimin Ren, Yan Liang, Baiqiao Su, Chaobo Sun, Hengtong Lu, Kaike Zhang, Chen Wei
Abstract
The ability of Large Language Models (LLMs) to precisely follow complex and fine-grained lexical instructions is a cornerstone of their utility and controllability. However, evaluating this capability remains a significant challenge. Current methods either rely on subjective and costly human evaluation or on automated ``LLM-as-a-judge
BibTeX
@inproceedings{aaai2026_lexinstructevall,
title = {LexInstructEval: Lexical Instruction Following Evaluation for Large Language Models},
author = {Huimin Ren and Yan Liang and Baiqiao Su and Chaobo Sun and Hengtong Lu and Kaike Zhang and Chen Wei},
booktitle = {AAAI 2026},
year = {2026}
}