← Search

Shuchen Wu

4 accepted papers

2026

Stay in Character, Stay Safe: Dual-Cycle Adversarial Self-Evolution for Role-Playing Agents

IJCAI 2026

LLM-based role-playing has rapidly improved in fidelity, yet stronger adherence to persona constraints commonly increases vulnerability to jailbreak attacks, especially for risky or negative personas. Most prior work mitigates this issue with training-time solutions (e.g., data curation or alignment

Cited by 0Scholar
2025

Building, Reusing, and Generalizing Abstract Representations from Concrete Sequences

ICLR 2025poster

Humans excel at learning abstract patterns across different sequences, filtering out irrelevant details, and transferring these generalized concepts to new sequences. In contrast, many sequence learning models lack the ability to abstract, which leads to memory inefficiency and poor transfer. We int…

Cited by 0SourcePDFScholar
2025

Concept-Guided Interpretability via Neural Chunking

NeurIPS 2025poster

Neural networks are often described as black boxes, reflecting the significant challenge of understanding their internal workings and interactions. We propose a different perspective that challenges the prevailing view: rather than being inscrutable, neural networks exhibit patterns in their raw po…

Cited by 0SourcecodeScholar
2022

Learning Structure from the Ground up---Hierarchical Representation Learning by Chunking

NeurIPS 2022accept

From learning to play the piano to speaking a new language, reusing and recombining previously acquired representations enables us to master complex skills and easily adapt to new environments. Inspired by the Gestalt principle of \textit{grouping by proximity} and theories of chunking in cognitive…

Cited by 17SourcePDFScholar