← Search

Tieru Wang

2 accepted papers

2026

Trajectory Generation with Conservative Value Guidance for Offline Reinforcement Learning

ICLR 2026poster

Recent advances in offline reinforcement learning (RL) have led to the development of high-performing algorithms that achieve impressive results across standard benchmarks. However, many of these methods depend on increasingly complex planning architectures, which hinder their deployment in real-wor…

Cited by 0SourceScholar