2025
Context-Aware Hierarchical Learning: A Two-Step Paradigm towards Safer LLMs
NeurIPS 2025poster
Large Language Models (LLMs) have emerged as powerful tools for diverse applications. However, their uniform token processing paradigm introduces critical vulnerabilities in instruction handling, particularly when exposed to adversarial scenarios. In this work, we identify and propose a novel class…