← Search

Jingyu Lei

2 accepted papers

2026

CORE: Compact Object-centric REpresentations as a New Paradigm for Token Merging in LVLMs

CVPR 2026

Large Vision-Language Models (LVLMs) usually suffer from prohibitive computational and memory costs due to the quadratic growth of visual tokens with image resolution. Existing token compression methods, while varied, often lack a high-level semantic understanding, leading to suboptimal merges, info

Cited by 0SourcecodeScholar
2025

RAPID: Recognition of Any-Possible DrIver Distraction via Multi-view Pose Generation Models

ICASSP 2025accepted

Driver distraction remains a pressing traffic safety issue. Drivers are often careless with their distraction behaviours, which may cause serious traffic accidents. However, current Driver Monitoring Systems (DMS) cannot be put into practical application well, which tend to have high latency, lack p…

Cited by 0SourceScholar