← Search

Jiafei Song

2 accepted papers

2026

EvoComp: Learning Visual Token Compression for Multimodal Large Language Models via Semantic-Guided Evolutionary Labeling

CVPR 2026

Recent Multimodal Large Language Models (MLLMs) have demonstrated strong performance on vision-language understanding tasks, yet their inference efficiency is often hampered by the large number of visual tokens, particularly in high-resolution or multi-image scenarios. To address this issue, we prop

Cited by 0SourceScholar
2020

Richer Aggregated Features for Optical Flow Estimation with Edge-aware Refinement

IROS 2020poster

Recent CNN-based optical flow approaches have a separated structure of feature extraction and flow estimation. The core task of optical flow is finding the corresponding points while rich representation is just the key part of such matching problems. However, the prior work usually pays more attenti…

Cited by 1SourceScholar