← Search

Ruijin Liu

4 accepted papers

2024

Align Before Adapt: Leveraging Entity-to-Region Alignments for Generalizable Video Action Recognition

CVPR 2024poster

Large-scale visual-language pre-trained models have achieved significant success in various video tasks. However most existing methods follow an "adapt then align" paradigm which adapts pre-trained image encoders to model video-level representations and utilizes one-hot or text embedding of the acti…

Cited by 9SourcePDFScholar
2022

Learning to Predict 3D Lane Shape and Camera Pose from a Single Image via Geometry Constraints

AAAI 2022technical

Detecting 3D lanes from the camera is a rising problem for autonomous vehicles. In this task, the correct camera pose is the key to generating accurate lanes, which can transform an image from perspective-view to the top-view. With this transformation, we can get rid of the perspective effects so th…

2021

Multimodal Transformer Networks for Pedestrian Trajectory Prediction

IJCAI 2021poster

We consider the problem of forecasting the future locations of pedestrians in an ego-centric view of a moving vehicle. Current CNNs or RNNs are flawed in capturing the high dynamics of motion between pedestrians and the ego-vehicle, and suffer from the massive parameter usages due to the inefficienc…