← Search

Zeyu Han

5 accepted papers

2026

Equi-RO: A 4D mmWave Radar Odometry via Equivariant Networks

RA-L 2026

Autonomous vehicles and robots rely on accurate odometry estimation in GPS-denied environments. While LiDARs and cameras struggle under extreme weather, 4D mmWave radar emerges as a robust alternative with all-weather operability and velocity measurement. In this paper, we introduce Equi-RO, an equi

Cited by 3SourceScholar
2026

Griffin: Aerial-Ground Cooperative Detection and Tracking Dataset and Benchmark

AAAI 2026technical

While cooperative perception can overcome the limitations of single-vehicle systems, the practical implementation of vehicle-to-vehicle and vehicle-to-infrastructure systems is often impeded by significant economic barriers. Aerial-ground cooperation (AGC), which pairs ground vehicles with drones, p

Cited by 0SourcePDFScholar
2025

A Generalized Control Revision Method for Autonomous Driving Safety

ICRA 2025

Safety is one of the most crucial challenges of autonomous driving vehicles, and one solution to guarantee safety is to employ an additional control revision module after the planning backbone. Control Barrier Function (CBF) has been widely used because of its strong mathematical foundation on safet

Cited by 0SourceScholar
2025

Rethinking Diffusion for Text-Driven Human Motion Generation: Redundant Representations, Evaluation, and Masked Autoregression

CVPR 2025poster

Since 2023, Vector Quantization (VQ)-based discrete generation methods have rapidly dominated human motion generation, primarily surpassing diffusion-based continuous generation methods in standard performance metrics. However, VQ-based methods have inherent limitations. Representing continuous moti…

Cited by 0SourcePDFScholar
2024

Zero-shot Referring Expression Comprehension via Structural Similarity Between Images and Captions

CVPR 2024poster

Zero-shot referring expression comprehension aims at localizing bounding boxes in an image corresponding to provided textual prompts which requires: (i) a fine-grained disentanglement of complex visual scene and textual context and (ii) a capacity to understand relationships among disentangled entit…