← Search

Tianyi Xiang

4 accepted papers

2025

Action Dubber: Timing Audible Actions via Inflectional Flow

ICML 2025poster

We introduce the task of Audible Action Temporal Localization, which aims to identify the spatio-temporal coordinates of audible movements. Unlike conventional tasks such as action recognition and temporal action localization, which broadly analyze video content, our task focuses on the distinct kin…

2025

Instance-Level Video Depth in Groups Beyond Occlusions

ICCV 2025poster

Depth estimation in dynamic, multi-object scenes remains a major challenge, especially under severe occlusions. Existing monocular models, including foundation models, struggle with instance-wise depth consistency due to their reliance on global regression. We tackle this problem from two key aspect…

Cited by 0SourcePDFScholar
2025

One-Shot Real-to-Sim via End-to-End Differentiable Simulation and Rendering

RA-L 2025

Identifying predictive world models for robots from sparse online observations is essential for robot task planning and execution in novel environments. However, existing methods that leverage differentiable programming to identify world models are incapable of jointly optimizing the geometry, appea

Cited by 5SourcecodeScholar
2022

Editing Out-of-Domain GAN Inversion via Differential Activations

ECCV 2022poster

"Despite the demonstrated editing capacity in the latent space of a pretrained GAN model, inverting real-world images is stuck in a dilemma that the reconstruction cannot be faithful to the original input. The main reason for this is that the distributions between training and real-world data are mi…