← Search

Max Wilcoxson

1 accepted papers

2025

Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration

ICML 2025poster

Unsupervised pretraining has been transformative in many supervised domains. However, applying such ideas to reinforcement learning (RL) presents a unique challenge in that fine-tuning does not involve mimicking task-specific data, but rather exploring and locating the solution through iterative sel…