2026
NExT-Guard: Training-Free Streaming Safeguard without Token-Level Labels
ICML 2026poster
Large language models are increasingly deployed in streaming scenarios, rendering conventional post-hoc safeguards ineffective as they fail to interdict unsafe content in real-time. While streaming safeguards based on token-level supervised training could address this, they necessitate expensive ann…