← Search

Varshanth Rao

2 accepted papers

2024

SCE-MAE: Selective Correspondence Enhancement with Masked Autoencoder for Self-Supervised Landmark Estimation

CVPR 2024poster

Self-supervised landmark estimation is a challenging task that demands the formation of locally distinct feature representations to identify sparse facial landmarks in the absence of annotated data. To tackle this task existing state-of-the-art (SOTA) methods (1) extract coarse features from backbon…

Cited by 1SourcePDFScholar
2022

Dual Perspective Network for Audio-Visual Event Localization

ECCV 2022poster

"The Audio-Visual Event Localization (AVEL) problem involves tackling three core sub-tasks: the creation of efficient audio-visual representations using cross-modal guidance, the formation of short-term temporal feature aggregations, and its accumulation to achieve long-term dependency resolution. T…

Cited by 21SourcePDFScholar