← Search

Xin Hao

5 accepted papers

2026

FedPDG: Prediction Discrepancy–Guided Data Generation for Heterogeneous Federated Learning

ICML 2026poster

One emerging approach to mitigating data heterogeneity in Federated Learning (FL) is to employ diffusion models to generate synthetic data for clients, thereby aligning local data distributions with the global distribution. Prior work has primarily focused on balance-oriented augmentation, which ass…

Cited by 0SourceScholar
2025

Explore the LiDAR-Camera Dynamic Adjustment Fusion for 3D Object Detection

ICRA 2025

Camera and LiDAR serve as informative sensors for accurate and robust autonomous driving systems. However, these sensors often exhibit heterogeneous natures, resulting in distributional modality gaps that present significant challenges for fusion. To address this, a robust fusion technique is crucia

Cited by 0SourcecodeScholar
2024

ViT-CoMer: Vision Transformer with Convolutional Multi-scale Feature Interaction for Dense Predictions

CVPR 2024highlight

Although Vision Transformer (ViT) has achieved significant success in computer vision it does not perform well in dense prediction tasks due to the lack of inner-patch information interaction and the limited diversity of feature scale. Most existing studies are devoted to designing vision-specific t…

2023

V2X-Seq: A Large-Scale Sequential Dataset for Vehicle-Infrastructure Cooperative Perception and Forecasting

CVPR 2023poster

Utilizing infrastructure and vehicle-side information to track and forecast the behaviors of surrounding traffic participants can significantly improve decision-making and safety in autonomous driving. However, the lack of real-world sequential datasets limits research in this area. To address this…

2021

Cross-Modality Person Re-Identification via Modality Confusion and Center Aggregation

ICCV 2021poster

Cross-modality person re-identification is a challenging task due to large cross-modality discrepancy and intra-modality variations. Currently, most existing methods focus on learning modality-specific or modality-shareable features by using the identity supervision or modality label. Different from…

Cited by 213PDFScholar