← Search

Alexandra Forsey-Smerek

1 accepted papers

2026

Masked IRL: LLM-Guided Reward Disambiguation from Demonstrations and Language

ICRA 2026poster

Robots can adapt to user preferences by learning reward functions from demonstrations, but with limited data, reward models often overfit to spurious correlations and fail to generalize. This happens because demonstrations show robots how to do a task but not what matters for that task, causing the …