← Search

Yanan Luo

2 accepted papers

2025

Video-Panda: Parameter-efficient Alignment for Encoder-free Video-Language Models

CVPR 2025poster

We present an efficient encoder-free approach for video-language understanding that achieves competitive performance while significantly reducing computational overhead. Current video-language models typically rely on heavyweight image encoders (300M-1.1B parameters) or video encoders (1B-1.4B param…

2022

Scene Consistency Representation Learning for Video Scene Segmentation

CVPR 2022poster

A long-term video, such as a movie or TV show, is composed of various scenes, each of which represents a series of shots sharing the same semantic story. Spotting the correct scene boundary from the long-term video is a challenging task, since a model must understand the storyline of the video to fi…

Cited by 20PDFcodeScholar