← Search

Junseo Jang

2 accepted papers

2025

Exploring the Impact of Instruction-Tuning on LLM’s Susceptibility to Misinformation

ACL 2025long

Instruction-tuning enhances the ability of large language models (LLMs) to follow user instructions more accurately, improving usability while reducing harmful outputs. However, this process may increase the model’s dependence on user input, potentially leading to the unfiltered acceptance of misinf…

Cited by 0SourcePDFScholar
2025

Small Changes, Big Impact: How Manipulating a Few Neurons Can Drastically Alter LLM Aggression

ACL 2025long

Recent remarkable advances in Large Language Models (LLMs) have led to innovations in various domains such as education, healthcare, and finance, while also raising serious concerns that they can be easily misused for malicious purposes. Most previous research has focused primarily on observing how…

Cited by 0SourcePDFScholar