2025
TrackVerse: A Large-Scale Object-Centric Video Dataset for Image-Level Representation Learning
ICCV 2025accepted
Video data inherently captures rich, dynamic contexts that reveal objects in varying poses, interactions, and state transitions, offering rich potential for unsupervised object representation learning. However, most prior representation learning methods rely on static image datasets like ImageNet, w…