← Search

In Cho

8 accepted papers

2026

4D Scaffold Gaussian Splatting with Dynamic-Aware Anchor Growing for Efficient and High-Fidelity Dynamic Scene Reconstruction

AAAI 2026technical

Modeling dynamic scenes through 4D Gaussians offers high visual fidelity and fast rendering speeds, but comes with significant storage overhead. Recent approaches mitigate this cost by aggressively reducing the number of Gaussians. However, this inevitably removes Gaussians essential for high-qualit

Cited by 10SourcePDFScholar
2026

Unsupervised Monocular 3D Keypoint Discovery from Multi-View Diffusion Priors

CVPR 2026

Most existing 3D keypoint estimation methods rely on manual annotations or calibrated multi-view images, both of which are expensive to collect.This paper introduces KeyDiff3D, a framework that can accurately predict 3D keypoints from a single image, thus eliminating the need for such expensive data

Cited by 0SourceScholar
2025

ExploreGS: Explorable 3D Scene Reconstruction with Virtual Camera Samplings and Diffusion Priors

ICCV 2025poster

Recent advances in novel view synthesis (NVS) have enabled real-time rendering with 3D Gaussian Splatting (3DGS). However, existing methods struggle with artifacts and missing regions when rendering unseen viewpoints, limiting seamless scene exploration. To address this, we propose a 3DGS-based pipe…

Cited by 0SourcePDFScholar
2025

Representing 3D Shapes with 64 Latent Vectors for 3D Diffusion Models

ICCV 2025poster

Constructing a compressed latent space through a variational autoencoder (VAE) is the key for efficient 3D diffusion models. This paper introduces COD-VAE that encodes 3D shapes into a COmpact set of 1D latent vectors without sacrificing quality. COD-VAE introduces a two-stage autoencoder scheme to…

2024

Hierarchically Structured Neural Bones for Reconstructing Animatable Objects from Casual Videos

ECCV 2024poster

"We propose a new framework for creating and easily manipulating 3D models of arbitrary objects using casually captured videos. Our core ingredient is a novel hierarchy deformation model, which captures motions of objects with a tree-structured bones. Our hierarchy system decomposes motions based on…

2024

Learning to Enhance Aperture Phasor Field for Non-Line-of-Sight Imaging

ECCV 2024poster

"This paper aims to facilitate more practical NLOS imaging by reducing the number of samplings and scan areas. To this end, we introduce a phasor-based enhancement network that is capable of predicting clean and full measurements from noisy partial observations. We leverage a denoising autoencoder s…

2019

Unsupervised Keypoint Learning for Guiding Class-Conditional Video Prediction

NeurIPS 2019poster

We propose a deep video prediction model conditioned on a single image and an action class. To generate future frames, we first detect keypoints of a moving object and predict future motion as a sequence of keypoints. The input image is then translated following the predicted keypoints sequence to c…

Cited by 58SourcePDFScholar