← Search

Qianzhun Wang

2 accepted papers

2026

E2HiL: Entropy-Guided Sample Selection for Efficient Real-World Human-in-the-Loop Reinforcement Learning

RA-L 2026

Human-in-the-loop guidance has emerged as an effective approach for accelerating online reinforcement learning (RL) in real-world manipulation. However, existing human-in-the-loop RL (HiL-RL) frameworks often suffer from low sample efficiency, requiring substantial human interventions to achieve con

Cited by 2SourceScholar
2025

SafeBimanual: Diffusion-based trajectory optimization for safe bimanual manipulation

CoRL 2025poster

Bimanual manipulation has been widely applied in household services and manufacturing, which enables the complex task completion with coordination requirements. Recent diffusion-based policy learning approaches have achieved promising performance in modeling action distributions for bimanual manipul…

Cited by 0SourceScholar