← Search

Ge Ren

3 accepted papers

2024

UOR: Universal Backdoor Attacks on Pre-trained Language Models

ACL 2024findings

Task-agnostic and transferable backdoors implanted in pre-trained language models (PLMs) pose a severe security threat as they can be inherited to any downstream task. However, existing methods rely on manual selection of triggers and backdoor representations, hindering their effectiveness and unive…

Cited by 20SourcePDFScholar
2021

Hate Speech Detection Based on Sentiment Knowledge Sharing

ACL 2021long

The wanton spread of hate speech on the internet brings great harm to society and families. It is urgent to establish and improve automatic detection and active avoidance mechanisms for hate speech. While there exist methods for hate speech detection, they stereotype words and hence suffer from inhe…