← Search

Chenhan Zhang

2 accepted papers

2026

Steering Beyond the Support: Adversarial Training on Unsupervised Jailbroken Activation Simulation

ICML 2026poster

Jailbreak prompts can trigger harmful comple- tions on aligned LLMs, In accordance, safety steering has been proposed: test-time activation interventions that steer jailbreak activations to trig- ger refusal while preserving benign utility. How- ever, existing steering methods are fundamentally supe…

Cited by 0SourceScholar
2025

EvaSR: Rethinking Efficient Visual Attention Design for Image Super-Resolution

ICASSP 2025accepted

Due to the advantages of long-range modeling via the self-attention mechanism, Transformer has taken various vision tasks by storm, including image super-resolution (SR). In this study, we reveal that the convolutional neural network (CNN) with proper visual attention is a more simple and effective…

Cited by 0SourceScholar