← Search

Zhixiong Nan

9 accepted papers

2026

AbductiveMLLM: Boosting Visual Abductive Reasoning Within MLLMs

AAAI 2026technical

Visual abductive reasoning (VAR) is a challenging task that requires AI systems to infer the most likely explanation for incomplete visual observations. While recent MLLMs develop strong general-purpose multimodal reasoning capabilities, they remain fall short in abductive inference, as compared to

Cited by 0SourcePDFScholar
2026

MUSE: Harnessing Precise and Diverse Semantics for Few-Shot Whole Slide Image Classification

CVPR 2026

In computational pathology, few-shot whole slide image classification is primarily driven by the extreme scarcity of expert-labeled slides. Recent vision-language methods incorporate textual semantics generated by large language models, but treat these descriptions as static class-level priors that

Cited by 0SourceScholar
2026

PGMamba: A Physical Model-Guided Global Mamba for Underwater Image Enhancement

AAAI 2026technical

Underwater image enhancement (UIE) aims to address image degradation caused by water absorption and scattering effects. Despite significant progress in deep learning-based UIE methods, existing approaches still face key challenges due to the neglect of physical imaging principle. Moreover, while cur

Cited by 0SourcePDFScholar
2025

MI-DETR: An Object Detection Model with Multi-time Inquiries Mechanism

CVPR 2025poster

Based on analyzing the character of cascaded decoder architecture commonly adopted in existing DETR-like models, this paper proposes a new decoder architecture. The cascaded decoder architecture constrains object queries to update in the cascaded direction, only enabling object queries to learn rela…

2024

DI-MaskDINO: A Joint Object Detection and Instance Segmentation Model

NeurIPS 2024poster

This paper is motivated by an interesting phenomenon: the performance of object detection lags behind that of instance segmentation (i.e., performance imbalance) when investigating the intermediate results from the beginning transformer decoder layer of MaskDINO (i.e., the SOTA model for joint detec…

Cited by 0SourcePDFScholar
2024

On-Road Object Importance Estimation: A New Dataset and A Model with Multi-Fold Top-Down Guidance

NeurIPS 2024poster

This paper addresses the problem of on-road object importance estimation, which utilizes video sequences captured from the driver's perspective as the input. Although this problem is significant for safer and smarter driving systems, the exploration of this problem remains limited. On one hand, publ…

Cited by 0SourcePDFScholar
2021

A Global-Local Coupling Two-Stage Path Planning Method for Mobile Robots

RA-L 2021

The path planning of mobile robots is an optimization problem that is difficult to solve directly owing to its nonlinear characteristics. This letter proposes the “global-local” Coupling Two-Stage Path Planning (CTSP) method. First, the globally optimal solution in the configuration space is given b

Cited by 53SourceScholar