← Search

Weihao Xia

8 accepted papers

2026

Quartet of Diffusions: Structure-Aware Point Cloud Generation through Part and Symmetry Guidance

ICLR 2026poster

We introduce the *Quartet of Diffusions*, a structure-aware point cloud generation framework that explicitly models part composition and symmetry. Unlike prior methods that treat shape generation as a holistic process or only support part composition, our approach leverages four coordinated diffusio…

Cited by 0SourceScholar
2026

Single-Stage fMRI-to-3D Reconstruction via Viewpoint-Aware Embedding and Hierarchical Guidance

AAAI 2026technical

Understanding the neural basis of three-dimensional (3D) perception is a fundamental objective in cognitive neuroscience. Despite advances in decoding 2D visual stimuli from neural data, reconstructing high-fidelity 3D objects with detailed texture and geometry remains largely unexplored. In this wo

Cited by 0SourcePDFScholar
2025

Temporal Streaming Batch Principal Component Analysis for Time Series Classification (Student Abstract)

AAAI 2025technical

In multivariate time series classification, although current sequence analysis models have excellent classification capabilities, they show significant shortcomings when dealing with long sequence multivariate data. This paper focuses on optimizing model performance for long-sequence multivariate da…

Cited by 0SourcePDFScholar
2022

High-Fidelity GAN Inversion with Padding Space

ECCV 2022poster

"Inverting a Generative Adversarial Network (GAN) facilitates a wide range of image editing tasks using pre-trained generators. Existing methods typically employ the latent space of GANs as the inversion space yet observe the insufficient recovery of spatial details. In this work, we propose to invo…

2022

Learning Quality-Aware Dynamic Memory for Video Object Segmentation

ECCV 2022poster

"Recently, several spatial-temporal memory-based methods have verified that storing intermediate frames and their masks as memory are helpful to segment target objects in videos. However, they mainly focus on better matching between the current frame and the memory frames without explicitly paying a…

2021

TediGAN: Text-Guided Diverse Face Image Generation and Manipulation

CVPR 2021poster

In this work, we propose TediGAN, a novel framework for multi-modal image generation and manipulation with textual descriptions. The proposed method consists of three components: StyleGAN inversion module, visual-linguistic similarity learning, and instance-level optimization. The inversion module m…

Cited by 472PDFcodeScholar