← Search

Wenyuan Xue

2 accepted papers

2024

Align Before Adapt: Leveraging Entity-to-Region Alignments for Generalizable Video Action Recognition

CVPR 2024poster

Large-scale visual-language pre-trained models have achieved significant success in various video tasks. However most existing methods follow an "adapt then align" paradigm which adapts pre-trained image encoders to model video-level representations and utilizes one-hot or text embedding of the acti…

Cited by 9SourcePDFScholar
2021

TGRNet: A Table Graph Reconstruction Network for Table Structure Recognition

ICCV 2021poster

A table arranging data in rows and columns is a very effective data structure, which has been widely used in business and scientific research. Considering large-scale tabular data in online and offline documents, automatic table recognition has attracted increasing attention from the document analys…

Cited by 69PDFcodeScholar