← Search

Heewoong Choi

2 accepted papers

2025

Option-aware Temporally Abstracted Value for Offline Goal-Conditioned Reinforcement Learning

NeurIPS 2025spotlight

Offline goal-conditioned reinforcement learning (GCRL) offers a practical learning paradigm in which goal-reaching policies are trained from abundant state–action trajectory datasets without additional environment interaction. However, offline GCRL still struggles with long-horizon tasks, even with…

Cited by 0SourcecodeScholar
2024

Listwise Reward Estimation for Offline Preference-based Reinforcement Learning

ICML 2024poster

In Reinforcement Learning (RL), designing precise reward functions remains to be a challenge, particularly when aligning with human intent. Preference-based RL (PbRL) was introduced to address this problem by learning reward models from human feedback. However, existing PbRL methods have limitations…