2025
Uni4D: Unifying Visual Foundation Models for 4D Modeling from a Single Video
CVPR 2025highlight
This paper presents a unified approach to understanding dynamic scenes from casual videos. Large pretrained vision foundation models, such as vision-language, video depth prediction, motion tracking, and segmentation models, offer promising capabilities. However, training a single model for comprehe…