Exploring the Impact of Instruction-Tuning on LLM’s Susceptibility to Misinformation
Instruction-tuning enhances the ability of large language models (LLMs) to follow user instructions more accurately, improving usability while reducing harmful outputs. However, this process may increase the model’s dependence on user input, potentially leading to the unfiltered acceptance of misinf…