2025
'What Did the Robot Do in My Absence?' Video Foundation Models to Enhance Intermittent Supervision
RA-L 2025
This paper investigates the use of Video Foundation Models (ViFMs) for generating robot data summaries to enhance intermittent human supervision of robot teams. We propose a novel framework that produces both generic and query-driven summaries of long-duration robot vision data in three modalities: