← Search

Zhihong Fu

3 accepted papers

2026

Scaling4D: Pushing the Frontier of Video Novel View Synthesis through Large-Scale Monocular Videos

CVPR 2026

Video Novel View Synthesis (VNVS) aims to render arbitrary novel viewpoints of dynamic scenes from a single-view video, but its algorithmic training faces a major challenge: the lack of large-scale multi-view video datasets. Prior methods often train on monocular data by framing it as an inpainting

Cited by 0SourceScholar
2022

SparseTT: Visual Tracking with Sparse Transformers

IJCAI 2022poster

Transformers have been successfully applied to the visual tracking task and significantly promote tracking performance. The self-attention mechanism designed to model long-range dependencies is the key to the success of Transformers. However, self-attention lacks focusing on the most relevant inform…

2021

STMTrack: Template-Free Visual Tracking With Space-Time Memory Networks

CVPR 2021poster

Boosting performance of the offline trained siamese trackers is getting harder nowadays since the fixed information of the template cropped from the first frame has been almost thoroughly mined, but they are poorly capable of resisting target appearance changes. Existing trackers with template updat…

Cited by 364PDFcodeScholar