← Search

Shouwei Ruan*

1 accepted papers

2024

Omniview-Tuning: Boosting Viewpoint Invariance of Vision-Language Pre-training Models

ECCV 2024oral

"Vision-Language Pre-training (VLP) models like CLIP have achieved remarkable success in computer vision and particularly demonstrated superior robustness to distribution shifts of 2D images. However, their robustness under 3D viewpoint variations is still limited, which can hinder the development f…