2026
Obedience or Vigilance? How Large Language Models React to Malicious Multiple-Choice Options (Student Abstract)
AAAI 2026technical
When evaluating large language models (LLMs) for question answering tasks, a common protocol is multiple-choice question-answering (MCQA), where the model selects from a fixed set of choices. In contemporary robustness testing, researchers typically perturb instructions or introduce confusion into f