← Search

Linjie Xu

4 accepted papers

2025

Efficient Discovery of Pareto Front for Multi-Objective Reinforcement Learning

ICLR 2025poster

Multi-objective reinforcement learning (MORL) excels at handling rapidly changing preferences in tasks that involve multiple criteria, even for unseen preferences. However, previous dominating MORL methods typically generate a fixed policy set or preference-conditioned policy through multiple traini…

Cited by 0SourcePDFScholar
2025

Unveiling Markov heads in Pretrained Language Models for Offline Reinforcement Learning

ICML 2025poster

Recently, incorporating knowledge from pretrained language models (PLMs) into decision transformers (DTs) has generated significant attention in offline reinforcement learning (RL). These PLMs perform well in RL tasks, raising an intriguing question: what kind of knowledge from PLMs has been transfe…

Cited by 0SourcePDFScholar
2024

Protecting Your LLMs with Information Bottleneck

NeurIPS 2024poster

The advent of large language models (LLMs) has revolutionized the field of natural language processing, yet they might be attacked to produce harmful content. Despite efforts to ethically align LLMs, these are often fragile and can be circumvented by jailbreaking attacks through optimized or manual…