← Search

Zhicheng He

4 accepted papers

2026

Granulon: Awakening Pixel-Level Visual Encoders with Adaptive Multi-Granularity Semantics for MLLM

CVPR 2026

Recent advances in multimodal large language models largely rely on CLIP-based visual encoders, which emphasize global semantic alignment but struggle with fine-grained visual understanding. In contrast, DINOv3 provides strong pixel-level perception yet lacks coarse-grained semantic abstraction, lea

Cited by 0SourcecodeScholar
2026

PolygMap: A Perceptive Locomotion Framework for Humanoid Robot Stair Climbing

ICRA 2026poster

Recently, biped robot walking technology has been significantly developed; however, mainly in a bland walking scheme. To emulate human walking, robots need to step on the positions they see in unknown spaces accurately. In this paper, we present PolyMap, a perception-based locomotion planning framew…

2024

CDM-MPC: An Integrated Dynamic Planning and Control Framework for Bipedal Robots Jumping

RA-L 2024

Performing acrobatic maneuvers like dynamic jumping in bipedal robots presents significant challenges in terms of actuation, motion planning, and control. Traditional approaches to these tasks often simplify dynamics to enhance computational efficiency, potentially overlooking critical factors such

Cited by 29SourceScholar
2023

A Survey on User Behavior Modeling in Recommender Systems

IJCAI 2023poster

User Behavior Modeling (UBM) plays a critical role in user interest learning, which has been extensively used in recommender systems. Crucial interactive patterns between users and items have been exploited, which brings compelling improvements in many recommendation tasks. In this paper, we attempt…

Cited by 35SourcePDFScholar