← Search

Haifeng Hu

9 accepted papers

2025

ARUNet: Advancing Real-Time Stereo Matching for Robotic Perception on Edge Devices

RA-L 2025

Accurate real-time 3D depth perception is crucial for robotic systems, enabling navigation, obstacle avoidance, and object manipulation. However, the high computational demands of stereo matching networks hinder their application on resource-constrained robotic platforms. This paper presents ARUNet,

Cited by 2SourceScholar
2025

CyIN: Cyclic Informative Latent Space for Bridging Complete and Incomplete Multimodal Learning

NeurIPS 2025poster

Multimodal machine learning, mimicking the human brain’s ability to integrate various modalities has seen rapid growth. Most previous multimodal models are trained on perfectly paired multimodal input to reach optimal performance. In real‑world deployments, however, the presence of modality is highl…

Cited by 0SourceScholar
2023

Maximum Entropy Population-Based Training for Zero-Shot Human-AI Coordination

AAAI 2023technical

We study the problem of training a Reinforcement Learning (RL) agent that is collaborative with humans without using human data. Although such agents can be obtained through self-play training, they can suffer significantly from the distributional shift when paired with unencountered partners, such…

2022

Multimodal Contrastive Learning via Uni-Modal Coding and Cross-Modal Prediction for Multimodal Sentiment Analysis

EMNLP 2022finding

Multimodal representation learning is a challenging task in which previous work mostly focus on either uni-modality pre-training or cross-modality fusion. In fact, we regard modeling multimodal representation as building a skyscraper, where laying stable foundation and designing the main structure a…

Cited by 22SourcePDFScholar
2021

Communicative Message Passing for Inductive Relation Reasoning

AAAI 2021technical

Relation prediction for knowledge graphs aims at predicting missing relationships between entities. Despite the importance of inductive relation prediction, most previous works are limited to a transductive setting and cannot process previously unseen entities. The recent proposed subgraph-based rel…

2021

Which is Making the Contribution: Modulating Unimodal and Cross-modal Dynamics for Multimodal Sentiment Analysis

EMNLP 2021finding

Multimodal sentiment analysis (MSA) draws increasing attention with the availability of multimodal data. The boost in performance of MSA models is mainly hindered by two problems. On the one hand, recent MSA works mostly focus on learning cross-modal dynamics, but neglect to explore an optimal solut…

Cited by 30SourcePDFScholar
2020

Adaptive Interaction Modeling via Graph Operations Search

CVPR 2020poster

Interaction modeling is important for video action analysis. Recently, several works design specific structures to model interactions in videos. However, their structures are manually designed and non-adaptive, which require structures design efforts and more importantly could not model interactions…

Cited by 7PDFcodeScholar