← Search

Xiaochen Yuan

7 accepted papers

2026

Cross-view Anchor Graph Learning and Factorization for Incomplete Multi-view Clustering

AAAI 2026technical

Graph-based incomplete multi-view clustering algorithms have gathered much attention due to their impressive clustering performance. However, existing methods primarily leverage intra-view correlation from observed views, while ignoring the exploration of explicit compensation relationships between

Cited by 0SourcePDFScholar
2026

FaceShield: Explainable Face Anti-Spoofing with Multimodal Large Language Models

AAAI 2026technical

Face anti-spoofing (FAS) is crucial for protecting facial recognition systems from presentation attacks. Previous methods approached this task as a classification problem, lacking interpretability and reasoning behind the predicted results. Recently, multimodal large language models (MLLMs) have sho

Cited by 0SourcePDFScholar
2026

PASA: Progressive-Adaptive Spectral Augmentation for Automated Auscultation in Data-Scarce Environments

AAAI 2026technical

Automated auscultation advances the detection of respiratory diseases, especially in areas with limited resources where traditional diagnostic methods are unavailable. On the other hand, the scarcity of auscultation datasets limits the automation performance, prompting the needs for data augmentatio

Cited by 0SourcePDFScholar
2026

SUGAR: Learning Skeleton Representation with Visual-Motion Knowledge for Action Recognition

AAAI 2026technical

Large Language Models (LLMs) hold rich implicit knowledge and powerful transferability. In this paper, we explore the combination of LLMs with the human skeleton to perform action classification and description. However, when treating LLM as a recognizer, two questions arise: 1) How can LLMs underst

Cited by 0SourcePDFScholar
2026

Spherical Procrustes Alignment for Reliable Medical Audio Diagnosis

ICML 2026poster

Reliable medical audio diagnosis demands models that are not only accurate but also honest about their uncertainty. However, fine-tuned models based on small, imbalanced datasets often become overconfident due to norm bias, whereby they rely on feature magnitude rather than semantic alignment. As a …

Cited by 0SourceScholar
2025

DADM: Dual Alignment of Domain and Modality for Face Anti-spoofing

ICCV 2025poster

With the availability of diverse sensor modalities (i.e., RGB, Depth, Infrared) and the success of multi-modal learning, multi-modal face anti-spoofing (FAS) has emerged as a prominent research focus. The intuition behind it is that leveraging multiple modalities can uncover more intrinsic spoofing…

2024

PVALane: Prior-Guided 3D Lane Detection with View-Agnostic Feature Alignment

AAAI 2024technical

Monocular 3D lane detection is essential for a reliable autonomous driving system and has recently been rapidly developing. Existing popular methods mainly employ a predefined 3D anchor for lane detection based on front-viewed (FV) space, aiming to mitigate the effects of view transformations. Howev…

Cited by 6SourcePDFScholar