← Search

Zhengxia Zou

13 accepted papers

2025

Unified Multi-Agent Trajectory Modeling with Masked Trajectory Diffusion

ICCV 2025poster

Understanding movements in multi-agent scenarios is a fundamental problem in intelligent systems. Previous research assumes complete and synchronized observations. However, real-world partial observation caused by occlusions leads to inevitable model failure, which demands a unified framework for co…

2023

Matting Moments: A Unified Data-Driven Matting Engine for Mobile AIGC in Photo Gallery

IJCAI 2023poster

Image matting is a fundamental technique in visual understanding and has become one of the most significant capabilities in mobile phones. Despite the development of mobile storage and computing power, achieving diverse mobile Artificial Intelligence Generated Content (AIGC) applications remains a g…

Cited by 3SourcePDFScholar
2023

Towards Unbiased Volume Rendering of Neural Implicit Surfaces With Geometry Priors

CVPR 2023poster

Learning surface by neural implicit rendering has been a promising way for multi-view reconstruction in recent years. Existing neural surface reconstruction methods, such as NeuS and VolSDF, can produce reliable meshes from multi-view posed images. Although they build a bridge between volume renderi…

2023

Zero-Shot Text-to-Parameter Translation for Game Character Auto-Creation

CVPR 2023poster

Recent popular Role-Playing Games (RPGs) saw the great success of character auto-creation systems. The bone-driven face model controlled by continuous parameters (like the position of bones) and discrete parameters (like the hairstyles) makes it possible for users to personalize and customize in-gam…

2022

A Unified Framework for Real Time Motion Completion

AAAI 2022technical

Motion completion, as a challenging and fundamental problem, is of great significance in film and game applications. For different motion completion application scenarios (in-betweening, in-filling, and blending), most previous methods deal with the completion problems with case-by-case methodology…

Cited by 23SourcePDFScholar
2022

Real-time Full-stack Traffic Scene Perception for Autonomous Driving with Roadside Cameras

ICRA 2022poster

We propose a novel and pragmatic framework for traffic scene perception with roadside cameras. The proposed framework covers a full-stack of roadside perception pipeline for infrastructure-assisted autonomous driving, including object detection, object localization, object tracking, and multi-camera…

Cited by 50SourceScholar
2021

Automatic Translation of Music-to-Dance for In-Game Characters

IJCAI 2021poster

Music-to-dance translation is an emerging and powerful feature in recent role-playing games. Previous works of this topic consider music-to-dance as a supervised motion generation problem based on time-series data. However, these methods require a large amount of training data pairs and may suffer f…

2021

Multi-View 3D Reconstruction With Transformers

ICCV 2021poster

Deep CNN-based methods have so far achieved the state of the art results in multi-view 3D object reconstruction. Despite the considerable progress, the two core modules of these methods - view feature extraction and multi-view fusion, are usually investigated separately, and the relations among mult…

Cited by 126PDFScholar
2020

Deep Adversarial Decomposition: A Unified Framework for Separating Superimposed Images

CVPR 2020poster

Separating individual image layers from a single mixed image has long been an important but challenging task. We propose a unified framework named "deep adversarial decomposition" for single superimposed image separation. Our method deals with both linear and non-linear mixtures under an adversarial…

Cited by 86PDFScholar
2019

Face-to-Parameter Translation for Game Character Auto-Creation

ICCV 2019poster

Character customization system is an important component in Role-Playing Games (RPGs), where players are allowed to edit the facial appearance of their in-game characters with their own preferences rather than using default templates. This paper proposes a method for automatically creating in-game c…

Cited by 75PDFScholar
2019

Generative Adversarial Training for Weakly Supervised Cloud Matting

ICCV 2019poster

The detection and removal of cloud in remote sensing images are essential for earth observation applications. Most previous methods consider cloud detection as a pixel-wise semantic segmentation process (cloud v.s. background), which inevitably leads to a category-ambiguity problem when dealing with…

Cited by 42PDFScholar