← Search

Anbang Wang

1 accepted papers

2026

Beyond Policy Training: Recursive Solution Search from Unannotated Videos

ICML 2026poster

Many real-world tasks are recorded as large collections of unannotated task executions, such as videos, which contain rich information about task progress but lack the supervision assumed by standard reinforcement learning (RL) pipelines. In many practical settings, the goal is not to train a reusab…

Cited by 0SourceScholar