2025
SecDecoding: Steerable Decoding for Safer LLM Generation
EMNLP 2025
Large language models (LLMs) have achieved remarkable performance across diverse tasks, yet ensuring output safety remains a fundamental challenge. Existing defense methods often suffer from limited generalization, high computational overhead, or significant utility degradation. In this work, we pre