← Search

Xudong Yu

2 accepted papers

2024

Contrastive Representation for Data Filtering in Cross-Domain Offline Reinforcement Learning

ICML 2024poster

Cross-domain offline reinforcement learning leverages source domain data with diverse transition dynamics to alleviate the data requirement for the target domain. However, simply merging the data of two domains leads to performance degradation due to the dynamics mismatch. Existing methods address t…

2024

Regularized Conditional Diffusion Model for Multi-Task Preference Alignment

NeurIPS 2024poster

Sequential decision-making can be formulated as a conditional generation process, with targets for alignment with human intents and versatility across various tasks. Previous return-conditioned diffusion models manifest comparable performance but rely on well-defined reward functions, which requires…

Cited by 7SourcePDFScholar