← Search

Dongyi Liu

3 accepted papers

2026

Mitigating the Safety Alignment Tax with Null-Space Constrained Policy Optimization

ICLR 2026poster

As Large Language Models (LLMs) are increasingly deployed in real-world applications, it is important to ensure their behaviors align with human values, societal norms, and ethical principles. However, safety alignment under Reinforcement Learning (RL) often suffers from forgetting learned general a…

Cited by 0SourcecodeScholar
2025

Attack by Yourself: Effective and Unnoticeable Multi-Category Graph Backdoor Attacks with Subgraph Triggers Pool

NeurIPS 2025poster

Graph Neural Networks (GNNs) have achieved significant success in various real-world applications, including social networks, finance systems, and traffic management. Recent researches highlight their vulnerability to backdoor attacks in node classification, where GNNs trained on a poisoned graph mi…

Cited by 0SourceScholar