← Search

Ruidi Chang

4 accepted papers

2025

Learning Distribution-wise Control in Representation Space for Language Models

ICML 2025poster

Interventions in language models (LMs) are applied strategically to steer model behavior during the forward pass. Learnable interventions, also known as representation fine-tuning, aim to apply pointwise control within the concept subspace and have proven effective in altering high-level behaviors.…

2025

Steering Information Utility in Key-Value Memory for Language Model Post-Training

NeurIPS 2025poster

Recent advancements in language models (LMs) have marked a shift toward the growing importance of post-training. Yet, post-training approaches such as supervised fine-tuning (SFT) do not guarantee the effective use of knowledge acquired during pretraining. We therefore introduce infosteer, a lightwe…

Cited by 0SourceScholar
2024

Large Language Model Based Multi-agents: A Survey of Progress and Challenges

IJCAI 2024poster

Large Language Models (LLMs) have achieved remarkable success across a wide array of tasks. Due to their notable capabilities in planning and reasoning, LLMs have been utilized as autonomous agents for the automatic execution of various tasks. Recently, LLM-based agent systems have rapidly evolved f…