← Search

Xiaochang Li

2 accepted papers

2026

Activation Steering for LLM Alignment via a Unified ODE-Based Framework

ICLR 2026poster

Activation steering, or representation engineering, offers a lightweight approach to align large language models (LLMs) by manipulating their internal activations at inference time. However, current methods suffer from two key limitations: \textit{(i)} the lack of a unified theoretical framework for…

Cited by 0SourcecodeScholar
2026

WestWorld: A Knowledge-Encoded Scalable Trajectory World Model for Diverse Robotic Systems

ICML 2026spotlight

Trajectory world models play a crucial role in robotic dynamics learning, planning, and control. While recent works have explored trajectory world models for diverse robotic systems, they struggle to scale to a large number of distinct system dynamics and overlook domain knowledge of physical struct…

Cited by 0SourcecodeScholar