← Search

Jiahui Liu

13 accepted papers

2026

ASSIST-3D: Adapted Scene Synthesis for Class-Agnostic 3D Instance Segmentation

AAAI 2026technical

Class-agnostic 3D instance segmentation tackles the challenging task of segmenting all object instances, including previously unseen ones, without semantic class reliance. Current methods struggle with generalization due to the scarce annotated 3D scene data or noisy 2D segmentations. While syntheti

Cited by 0SourcePDFScholar
2026

From Winning to Understanding: A Diagnostic Long-Horizon RTS Benchmark for LLMs

ICML 2026poster

Large language models (LLMs) are increasingly used as decision modules, yet existing benchmarks provide limited coverage of long-horizon, adversarial interaction while faithfully acting on human instructions. We introduce a long-horizon Red Alert RTS benchmark with a hierarchical interface in which …

Cited by 0SourceScholar
2026

Learning to See through Illumination Extremes with Event Streaming in Multimodal Large Language Models

CVPR 2026

Multimodal Large Language Models (MLLMs) perform strong vision-language reasoning under standard conditions but fail in extreme illumination, where RGB inputs lose irrevocable structure and semantics. We propose Event-MLLM, an event-enhanced model that performs all-light visual reasoning by dynamica

Cited by 0SourceScholar
2025

Aligning Effective Tokens with Video Anomaly in Large Language Models

ICCV 2025poster

Understanding abnormal events in videos is a vital and challenging task that has garnered significant attention in a wide range of applications. Although current video understanding Multi-modal Large Language Models (MLLMs) are capable of analyzing general videos, they often struggle to handle anoma…

Cited by 0SourcePDFScholar
2025

Dual-BEV Nav: Dual-Layer BEV-Based Heuristic Path Planning for Robotic Navigation in Unstructured Outdoor Environments

ICRA 2025

Path planning with strong environmental adaptability plays a crucial role in robotic navigation in unstructured outdoor environments, especially in the case of low-quality location and map information. The path planning ability of a robot depends on the identification of the traversability of global

Cited by 2SourceScholar
2025

Equipping Vision Foundation Model with Mixture of Experts for Out-of-Distribution Detection

ICCV 2025poster

Pre-trained vision foundation models have transformed many computer vision tasks. Despite their strong ability to learn discriminative and generalizable features crucial for out-of-distribution (OOD) detection, their impact on this task remains underexplored. Motivated by this gap, we systematically…

Cited by 0SourcePDFScholar
2025

How Far are AI-generated Videos from Simulating the 3D Visual World: A Learned 3D Evaluation Approach

ICCV 2025poster

Recent advancements in video diffusion models enable the generation of photorealistic videos with impressive 3D consistency and temporal coherence. However, the extent to which these AI-generated videos simulate the 3D visual world remains underexplored. In this paper, we introduce Learned 3D Evalua…

Cited by 0SourcePDFScholar
2025

Learning from Neighbors: Category Extrapolation for Long-Tail Learning

CVPR 2025poster

Balancing training on long-tail data distributions remains a long-standing challenge in deep learning. While methods such as re-weighting and re-sampling help alleviate the imbalance issue, limited sample diversity continues to hinder models from learning robust and generalizable feature representat…

Cited by 0SourcePDFScholar
2024

Compositional Generalization for Multi-Label Text Classification: A Data-Augmentation Approach

AAAI 2024technical

Despite significant advancements in multi-label text classification, the ability of existing models to generalize to novel and seldom-encountered complex concepts, which are compositions of elementary ones, remains underexplored. This research addresses this gap. By creating unique data splits acros…

2023

GICI-LIB: A GNSS/INS/Camera Integrated Navigation Library

RA-L 2023

Accurate navigation is essential for autonomous robots and vehicles. In recent years, the integration of the Global Navigation Satellite System (GNSS), Inertial Navigation System (INS), and camera has garnered considerable attention due to its robustness and high accuracy in diverse environments. Ho

Cited by 39SourcecodeScholar
2023

MarS3D: A Plug-and-Play Motion-Aware Model for Semantic Segmentation on Multi-Scan 3D Point Clouds

CVPR 2023poster

3D semantic segmentation on multi-scan large-scale point clouds plays an important role in autonomous systems. Unlike the single-scan-based semantic segmentation task, this task requires distinguishing the motion states of points in addition to their semantic categories. However, methods designed fo…