← Search

Shaopeng Yang

3 accepted papers

2025

Bridging Gait Recognition and Large Language Models Sequence Modeling

CVPR 2025poster

Gait sequences exhibit sequential structures and contextual relationships similar to those in natural language, where each element--whether a word or a gait step--is connected to its predecessors and successors. This similarity enables the transformation of gait sequences into "texts" containing ide…

Cited by 1SourcePDFScholar
2024

RoScenes: A Large-scale Multi-view 3D Dataset for Roadside Perception

ECCV 2024poster

"We introduce RoScenes, the largest multi-view roadside perception dataset, which aims to shed light on the development of vision-centric Bird’s Eye View (BEV) approaches for more challenging traffic scenes. The highlights of RoScenes include significantly large perception area, full scene coverage…

2022

CrowdFormer: An Overlap Patching Vision Transformer for Top-Down Crowd Counting

IJCAI 2022poster

Crowd counting methods typically predict a density map as an intermediate representation of counting, and achieve good performance. However, due to the perspective phenomenon, there is a scale variation in real scenes, which causes the density map-based methods suffer from a severe scene generalizat…