← Search

Haiyang Wu

8 accepted papers

2026

Gaussian or Plane? Both: Semantic-Driven Voxel Representation for LiDAR-Inertial Odometry

RA-L 2026

Accurate LiDAR-inertial odometry (LIO) highly depends on the geometric fidelity of the underlying environment representation. We explore the new and interesting research direction of integrating semantic segmentation models into metric odometry algorithms to enrich their representational capacity. S

Cited by 2SourceScholar
2026

Gaussian or Plane? Both: Semantic-Driven Voxel Representation for LiDAR–Inertial Odometry

ICRA 2026poster

Accurate LiDAR-inertial odometry (LIO) highly depends on the geometric fidelity of the underlying environment representation. We explore the new and interesting research direction of integrating semantic segmentation models into metric odometry algorithms to enrich their representational capacity. S…

Cited by 0SourceScholar
2026

Multi-Agent VLMs Guided Self-Training with PNU Loss for Low-Resource Offensive Content Detection

AAAI 2026technical

Accurate detection of offensive content on social media demands high-quality labeled data; however, such data is often scarce due to the low prevalence of offensive instances and the high cost of manual annotation. To address this low-resource challenge, we propose a self-training framework that lev

Cited by 0SourcePDFScholar
2025

Behavior Importance-Aware Graph Neural Architecture Search for Cross-Domain Recommendation

AAAI 2025technical

Cross-domain recommendation (CDR) mitigates data sparsity and cold-start issues in recommendation systems. While recent CDR approaches using graph neural networks (GNNs) capture complex user-item interactions, they rely on manually designed architectures that are often suboptimal and labor-intensive…

2025

CPCF: A Cross-Prompt Contrastive Framework for Referring Multimodal Large Language Models

ICML 2025poster

Referring MLLMs extend conventional multimodal large language models by allowing them to receive referring visual prompts and generate responses tailored to the indicated regions. However, these models often suffer from suboptimal performance due to incorrect responses tailored to misleading areas a…

Cited by 0SourcePDFScholar
2025

LOHRec: Leveraging Order and Hierarchy in Generative Sequential Recommendation

EMNLP 2025

The sequential recommendation task involves predicting the items users will be interested in next based on their past interaction sequence. Recently, sequential recommender systems with generative retrieval have garnered significant attention. However, during training, these generative recommenders

2025

POPEN: Preference-Based Optimization and Ensemble for LVLM-Based Reasoning Segmentation

CVPR 2025poster

Existing LVLM-based reasoning segmentation methods often suffer from imprecise segmentation results and hallucinations in their text responses. This paper introduces POPEN, a novel framework designed to address these issues and achieve improved results. POPEN includes a preference-based optimization…

Cited by 2SourcePDFScholar
2025

Retrv-R1: A Reasoning-Driven MLLM Framework for Universal and Efficient Multimodal Retrieval

NeurIPS 2025poster

The success of DeepSeek-R1 demonstrates the immense potential of using reinforcement learning (RL) to enhance LLMs' reasoning capabilities. This paper introduces Retrv-R1, the first R1-style MLLM specifically designed for multimodal universal retrieval, achieving higher performance by employing step…

Cited by 0SourceScholar