← Search

Juan Chen

3 accepted papers

2025

Glance2Gaze: Efficient Vision-Language Models from Glance Fusion to Gaze Compression

NeurIPS 2025poster

Vision-language models heavily rely on visual representations, yet ensuring its efficiency remains a critical challenge. Most existing approaches focus on reducing visual tokens either at the visual encoder phase or during the LLM decoder stage. Inspired by human visual cognition, where an initial g…

Cited by 0SourceScholar
2024

RVDNet: A Two-Stage Network for Real-World Video Desnowing with Domain Adaptation

ICASSP 2024accepted

Video snow removal is an important task in computer vision, as the snowflakes in videos reduce visibility and negatively affect the performance of outdoor visual systems. However, due to the complexity of real snowy scenarios, it is difficult to apply existing supervised learning-based methods to pr…

Cited by 0SourceScholar
2021

Autonomous Navigation for Adaptive Unmanned Underwater Vehicles Using Fiducial Markers

ICRA 2021poster

This paper presents an integrated methodology and experimental validation of an autonomous framework for unmanned underwater vehicles (UUVs) merely equipped with a conventional monocular camera and a pressure sensor to accomplish high-performance autonomy. Optimal pose of the UUV is solved iterative…

Cited by 24SourceScholar