2023
Explanation-based Finetuning Makes Models More Robust to Spurious Cues
ACL 2023long
Large Language Models (LLMs) are so powerful that they sometimes learn correlations between labels and features that are irrelevant to the task, leading to poor generalization on out-of-distribution data. We propose explanation-based finetuning as a general approach to mitigate LLMs’ reliance on spu…