← Search

Changxi Zheng

16 accepted papers

2026

Real-To-Sim Robot Policy Evaluation with Gaussian Splatting Simulation of Soft-Body Interactions

ICRA 2026poster

Robotic manipulation policies are advancing rapidly, but their direct evaluation in the real world remains costly, time-consuming, and difficult to reproduce, particularly for tasks involving deformable objects. Simulation provides a scalable and systematic alternative, yet existing simulators often…

2025

CAT4D: Create Anything in 4D with Multi-View Video Diffusion Models

CVPR 2025poster

We present CAT4D, a method for creating 4D (dynamic 3D) scenes from monocular video. CAT4D leverages a multi-view video diffusion model trained on a diverse combination of datasets to enable novel view synthesis at any specified camera poses and timestamps. Combined with a novel sampling approach, t…

2025

VLMaterial: Procedural Material Generation with Large Vision-Language Models

ICLR 2025spotlight

Procedural materials, represented as functional node graphs, are ubiquitous in computer graphics for photorealistic material appearance design. They allow users to perform intuitive and precise editing to achieve desired visual appearances. However, creating a procedural material given an input imag…

Cited by 0SourcePDFScholar
2024

Generative Camera Dolly: Extreme Monocular Dynamic Novel View Synthesis

ECCV 2024oral

"Accurate reconstruction of complex dynamic scenes from just a single viewpoint continues to be a challenging task in computer vision. Current dynamic novel view synthesis methods typically require videos from many different camera viewpoints, necessitating careful recording setups, and significantl…

Cited by 23SourcePDFScholar
2024

Physics-Based Interaction with 3D Objects via Video Generation

ECCV 2024oral

"Realistic object interactions are crucial for creating immersive virtual experiences, yet synthesizing realistic 3D object dynamics in response to novel interactions remains a significant challenge. Unlike unconditional or text-conditioned dynamics generation, action-conditioned dynamics requires p…

2024

Sin3DM: Learning a Diffusion Model from a Single 3D Textured Shape

ICLR 2024poster

Synthesizing novel 3D models that resemble the input example as long been pursued by graphics artists and machine learning researchers. In this paper, we present Sin3DM, a diffusion model that learns the internal patch distribution from a single 3D textured shape and generates high-quality variation…

2023

Implicit Neural Spatial Representations for Time-dependent PDEs

ICML 2023poster

Implicit Neural Spatial Representation (INSR) has emerged as an effective representation of spatially-dependent vector fields. This work explores solving time-dependent PDEs with INSR. Classical PDE solvers introduce both temporal and spatial discretizations. Common spatial discretizations include m…

Cited by 33SourcePDFScholar
2022

Dynamic Sliding Window for Realtime Denoising Networks

ICASSP 2022accepted

Realtime speech denoising has been long studied. Almost all existing methods process the incoming data stream using a sliding window of fixed-size. Yet, we show that the use of fixed-size sliding window may lead to an accumulating lag, especially in presence of other background computing processes t…

Cited by 0SourceScholar
2022

FishGym: A High-Performance Physics-based Simulation Framework for Underwater Robot Learning

ICRA 2022poster

Bionic underwater robots have demonstrated their superiority in many applications. Yet, training their intelligence for a variety of tasks that mimic the behavior of underwater creatures poses a number of challenges in practice, mainly due to lack of a large amount of available training data as well…

Cited by 14SourceScholar
2020

Listening to Sounds of Silence for Speech Denoising

NeurIPS 2020poster

We introduce a deep learning model for speech denoising, a long-standing challenge in audio analysis arising in numerous applications. Our approach is based on a key observation about human speech: there is often a short pause between each sentence or word. In a recorded speech signal, those pauses…

2020

One Man's Trash Is Another Man's Treasure: Resisting Adversarial Examples by Adversarial Examples

CVPR 2020poster

Modern image classification systems are often built on deep neural networks, which suffer from adversarial examples--images with deliberately crafted, imperceptible noise to mislead the network's classification. To defend against adversarial examples, a plausible idea is to obfuscate the network's g…

Cited by 31PDFcodeScholar
2019

Rethinking Generative Mode Coverage: A Pointwise Guaranteed Approach

NeurIPS 2019poster

Many generative models have to combat missing modes. The conventional wisdom to this end is by reducing through training a statistical distance (such as f -divergence) between the generated distribution and provided data distribution. But this is more of a heuristic than a guarantee. The statistical…

Cited by 26SourcePDFScholar