← Search

Tu Ouyang

1 accepted papers

2026

EASE: Practical and Efficient Safety Alignment for Small Language Models

AAAI 2026technical

Small language models (SLMs) are increasingly deployed on edge devices, making their safety alignment crucial yet challenging. Current shallow alignment methods that rely on direct refusal of malicious queries fail to provide robust protection, particularly against adversarial jailbreaks. While deli

Cited by 0SourcePDFScholar