← Search

Claas Voelcker

3 accepted papers

2026

Test-Time Graph Search for Goal-Conditioned Reinforcement Learning

ICML 2026poster

Offline goal-conditioned reinforcement learning (GCRL) often struggles with long-horizon tasks, where errors in value estimation accumulate and produce unreliable policies. It is typically assumed that effective long-term planning is infeasible without specialized training. In contrast, our work dem…

Cited by 0SourceScholar
2026

Trust-Region Diffusion Policies for Massively Parallel On-Policy RL

ICML 2026poster

Reinforcement learning with massively parallel simulations has become an emerging trend; however, most existing approaches still rely on simple Gaussian policy parameterizations. Diffusion models provide a more expressive policy class and have shown strong performance on challenging control problems…

Cited by 0SourceScholar
2020

Structured Object-Aware Physics Prediction for Video Modeling and Planning

ICLR 2020poster

When humans observe a physical system, they can easily locate components, understand their interactions, and anticipate future behavior, even in settings with complicated and previously unseen interactions. For computers, however, learning such models from videos in an unsupervised fashion is an uns…

Cited by 73SourcecodeScholar