← Search

Jayce Haoran Wang

1 accepted papers

2025

Efficient Imitation Without Demonstrations via Value-Penalized Auxiliary Control from Examples

ICRA 2025

Common approaches to providing feedback in reinforcement learning are the use of hand-crafted rewards or full-trajectory expert demonstrations. Alternatively, one can use examples of completed tasks, but such an approach can be extremely sample inefficient. We introduce value-penalized auxiliary con

Cited by 0SourcecodeScholar