← Search

Yuexi Zhang

2 accepted papers

2024

HAT: History-Augmented Anchor Transformer for Online Temporal Action Localization

ECCV 2024poster

"Online video understanding often relies on individual frames, leading to frame-by-frame predictions. Recent advancements such as Online Temporal Action Localization (OnTAL), extend this approach to instance-level predictions. However, existing methods mainly focus on short-term context, neglecting…

2020

Key Frame Proposal Network for Efficient Pose Estimation in Videos

ECCV 2020poster

Human pose estimation in video relies on local information by either estimating each frame independently or tracking poses across frames. In this paper, we propose a novel method combining local approaches with global context. We introduce a light weighted, unsupervised, key-frame proposal network (…