← Search

Gang Tan

1 accepted papers

2026

Chain-of-Thought Driven Adversarial Scenario Extrapolation for Robust Language Models

AAAI 2026technical

Large Language Models (LLMs) exhibit impressive capabilities, but remain susceptible to a growing spectrum of safety risks, including jailbreaks, toxic content, hallucinations, and bias. Existing defenses often address only a single threat type or resort to rigid outright rejection, sacrificing user

Cited by 0SourcePDFScholar