← Search

Yow-Fu Liou

1 accepted papers

2026

Obedience or Vigilance? How Large Language Models React to Malicious Multiple-Choice Options (Student Abstract)

AAAI 2026technical

When evaluating large language models (LLMs) for question answering tasks, a common protocol is multiple-choice question-answering (MCQA), where the model selects from a fixed set of choices. In contemporary robustness testing, researchers typically perturb instructions or introduce confusion into f

Cited by 0SourcePDFScholar