← Search

Yujian Wei

1 accepted papers

2024

Making Harmful Behaviors Unlearnable for Large Language Models

ACL 2024findings

Large language models (LLMs) have shown great potential to empower various domains and are often customized by fine-tuning for the requirements of different applications. However, the powerful learning ability of LLMs not only enables them to learn new tasks but also makes them vulnerable to learnin…