2024
Rethinking Pragmatics in Large Language Models: Towards Open-Ended Evaluation and Preference Tuning
EMNLP 2024main
This study addresses the challenges of assessing and enhancing social-pragmatic inference in large language models (LLMs). We first highlight the inadequacy of current accuracy-based multiple choice question answering (MCQA) formats in assessing social-pragmatic reasoning, and propose the direct eva…