← Search

Kuan Heng Lin

3 accepted papers

2026

Vista4D: Video Reshooting with 4D Point Clouds

CVPR 2026

We present **Vista4D**, a robust and flexible video reshooting framework that grounds the input video and target cameras in a 4D point cloud. Specifically, given an input video, our method re-synthesizes the scene with the same dynamics from a different camera trajectory and viewpoint. Existing vide

Cited by 0SourcecodeScholar
2024

Ctrl-X: Controlling Structure and Appearance for Text-To-Image Generation Without Guidance

NeurIPS 2024poster

Recent controllable generation approaches such as FreeControl and Diffusion Self-Guidance bring fine-grained spatial and appearance control to text-to-image (T2I) diffusion models without training auxiliary modules. However, these methods optimize the latent embedding for each type of score function…

2024

FreeControl: Training-Free Spatial Control of Any Text-to-Image Diffusion Model with Any Condition

CVPR 2024poster

Recent approaches such as ControlNet offer users fine-grained spatial control over text-to-image (T2I) diffusion models. However auxiliary modules have to be trained for each spatial condition type model architecture and checkpoint putting them at odds with the diverse intents and preferences a huma…