2026
Efficient Video Object Segmentation and Tracking with Recurrent Dynamic Submodel
CVPR 2026
Large vision foundation models, such as SAM2, have achieved remarkable performance in video object segmentation and tracking (VOST). However, their effectiveness is hindered by significant computational overhead. While model pruning is a widely used strategy to address this issue, traditional static