← Search

Tien-Tsin Wong

15 accepted papers

2026

Learning to Control Physically-simulated 3D Characters via Generating and Mimicking 2D Motions

CVPR 2026

Video data is more cost-effective than motion capture data for learning 3D character controllers, yet using it to generate realistic and physically plausible motions remains challenging. Previous approaches typically rely on off-the-shelf motion reconstruction techniques to extract 3D kinematic traj

Cited by 0SourcecodeScholar
2026

Mamba-Driven Multi-View Discriminative Clustering via Global-Local Cross-View Sequence Modeling

AAAI 2026technical

Multi-view clustering (MVC) has recently garnered increasing attention for its ability to partition unlabeled samples into distinct clusters by leveraging complementary and consistent information from different views. Existing MVC methods primarily combine deep neural networks with contrastive learn

Cited by 0SourcePDFScholar
2025

Advancing Manga Analysis: Comprehensive Segmentation Annotations for the Manga109 Dataset

CVPR 2025poster

Manga, a popular form of multimodal artwork, has traditionally been overlooked in deep learning advancements due to the absence of a robust dataset and comprehensive annotation. Manga segmentation is the key to the digital migration of manga. There exists a significant domain gap between the manga a…

Cited by 0SourcePDFScholar
2025

BlueNeg: A 35mm Negative Film Dataset for Restoring Channel-Heterogeneous Deterioration

ICCV 2025poster

While digitally acquired photographs have been dominating since around 2000, there remains a huge amount of legacy photographs being acquired by optical cameras and are stored in the form of film negatives. In this paper, we address the unique challenge of channel-heterogeneous deterioration in film…

2025

Consistent and Controllable Image Animation with Motion Diffusion Models

CVPR 2025poster

Diffusion models have achieved significant progress in the task of image animation due to their powerful generative capabilities. However, preserving appearance consistency to the static input image, and avoiding abrupt motion change in the generated animation, remains challenging. In this paper, we…

Cited by 0SourcePDFScholar
2025

VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

ICCV 2025accepted

Video diffusion models (VDMs) have advanced significantly in recent years, enabling the generation of highly realistic videos and drawing the attention of the community in their potential as world simulators. However, despite their capabilities, VDMs often fail to produce physically plausible videos…

2024

DynamiCrafter: Animating Open-domain Images with Video Diffusion Priors

ECCV 2024oral

"Animating a still image offers an engaging visual experience. Traditional image animation techniques mainly focus on animating natural scenes with stochastic dynamics (e.g. clouds and fluid) or domain-specific motions (e.g. human hair or body motions), and thus limits their applicability to more ge…

2023

CodeTalker: Speech-Driven 3D Facial Animation With Discrete Motion Prior

CVPR 2023poster

Speech-driven 3D facial animation has been widely studied, yet there is still a gap to achieving realism and vividness due to the highly ill-posed nature and scarcity of audio-visual data. Existing works typically formulate the cross-modal mapping into a regression task, which suffers from the regre…

2021

Bidirectional Projection Network for Cross Dimension Scene Understanding

CVPR 2021poster

2D image representations are in regular grids and can be processed efficiently, whereas 3D point clouds are unordered and scattered in 3D space. The information inside these two visual domains is well complementary, e.g., 2D images have fine-grained texture while 3D point clouds contain plentiful ge…

Cited by 147PDFcodeScholar
2021

User-Guided Line Art Flat Filling With Split Filling Mechanism

CVPR 2021poster

Flat filling is a critical step in digital artistic content creation with the objective of filling line arts with flat colors. We present a deep learning framework for user-guided line art flat filling that can compute the "influence areas" of the user color scribbles, i.e., the areas where the user…

Cited by 77PDFScholar
2020

Erasing Appearance Preservation in Optimization-based Smoothing

ECCV 2020poster

Optimization-based image smoothing is routinely formulated as the game between a smoothing energy and an appearance preservation energy. Achieving adequate smoothing is a fundamental goal of these image smoothing algorithms. We show that partially ""erasing"" the appearance preservation facilitate a…

Cited by 8SourcePDFScholar