← Search

Seung-Hwan Baek

27 accepted papers

2026

Learning to Generate Highly Dynamic Videos using Synthetic Motion Data

CVPR 2026

Despite recent progress, video diffusion models still struggle to synthesize realistic videos involving highly dynamic motions or requiring fine-grained motion controllability. A central limitation lies in the scarcity of such examples in commonly used training datasets. To address this, we introduc

Cited by 0SourceScholar
2025

A Real-world Display Inverse Rendering Dataset

ICCV 2025poster

Inverse rendering aims to reconstruct geometry and reflectance from captured images. Display-camera imaging systems offer unique advantages for this task: each pixel can easily function as a programmable point light source, and the polarized light emitted by LCD displays facilitates diffuse-specular…

2025

Dense Dispersed Structured Light for Hyperspectral 3D Imaging of Dynamic Scenes

CVPR 2025poster

Hyperspectral 3D imaging captures both depth maps and hyperspectral images, enabling comprehensive geometric and material analysis. Recent methods achieve high spectral and depth accuracy; however, they require long acquisition times--often over several minutes--or rely on large, expensive systems,…

Cited by 0SourcePDFScholar
2025

Dual Exposure Stereo for Extended Dynamic Range 3D Imaging

CVPR 2025poster

Achieving robust stereo 3D imaging under diverse illumination conditions is an important however challenging task, largely due to the limited dynamic ranges (DRs) of cameras, which are significantly smaller than real world DR. As a result, the accuracy of existing stereo depth estimation methods is…

Cited by 0SourcePDFScholar
2025

FloVD: Optical Flow Meets Video Diffusion Model for Enhanced Camera-Controlled Video Synthesis

CVPR 2025poster

We present FloVD, a novel video diffusion model for camera-controllable video generation. FloVD leverages optical flow to represent the motions of the camera and moving objects. This approach offers two key benefits. Since optical flow can be directly estimated from videos, our approach allows for t…

Cited by 6SourcePDFScholar
2024

ParamISP: Learned Forward and Inverse ISPs using Camera Parameters

CVPR 2024poster

RAW images are rarely shared mainly due to its excessive data size compared to their sRGB counterparts obtained by camera ISPs. Learning the forward and inverse processes of camera ISPs has been recently demonstrated enabling physically-meaningful RAW-level image processing on input sRGB images. How…

2024

Polarization Wavefront Lidar: Learning Large Scene Reconstruction from Polarized Wavefronts

CVPR 2024poster

Lidar has become a cornerstone sensing modality for 3D vision especially for large outdoor scenarios and autonomous driving. Conventional lidar sensors are capable of providing centimeter-accurate distance information by emitting laser pulses into a scene and measuring the time-of-flight (ToF) of th…

Cited by 1SourcePDFScholar
2024

Spectral and Polarization Vision: Spectro-polarimetric Real-world Dataset

CVPR 2024highlight

Image datasets are essential not only in validating existing methods in computer vision but also in developing new methods. Many image datasets exist consisting of trichromatic intensity images taken with RGB cameras which are designed to replicate human vision. However polarization and spectrum the…

Cited by 3SourcePDFScholar
2023

Multi-view Spectral Polarization Propagation for Video Glass Segmentation

ICCV 2023poster

In this paper, we present the first polarization-guided video glass segmentation propagation solution (PGVS-Net) that can robustly and coherently propagate glass segmentation in RGB-P video sequences. By leveraging spatiotemporal polarization and color information, our method combines multi-view pol…

Cited by 8PDFScholar
2023

Polarimetric iToF: Measuring High-Fidelity Depth Through Scattering Media

CVPR 2023highlight

Indirect time-of-flight (iToF) imaging allows us to capture dense depth information at a low cost. However, iToF imaging often suffers from multipath interference (MPI) artifacts in the presence of scattering media, resulting in severe depth-accuracy degradation. For instance, iToF cameras cannot me…

Cited by 8SourcePDFScholar
2022

BigColor: Colorization Using a Generative Color Prior for Natural Images

ECCV 2022poster

"For realistic and vivid colorization, generative priors have recently been exploited. However, such generative priors often fail for in-the-wild complex images due to their limited representation space. In this paper, we propose BigColor, a novel colorization approach that provides vivid colorizati…

2022

Glass Segmentation Using Intensity and Spectral Polarization Cues

CVPR 2022poster

Transparent and semi-transparent materials pose significant challenges for existing scene understanding and segmentation algorithms due to their lack of RGB texture which impedes the extraction of meaningful features. In this work, we exploit that the light-matter interactions on glass materials pro…

Cited by 93PDFScholar
2021

Mask-ToF: Learning Microlens Masks for Flying Pixel Correction in Time-of-Flight Imaging

CVPR 2021poster

We introduce Mask-ToF, a method to reduce flying pixels (FP) in time-of-flight (ToF) depth captures. FPs are pervasive artifacts which occur around depth edges, where light paths from both an object and its background are integrated over the aperture. This light mixes at a sensor pixel to produce er…

Cited by 23PDFcodeScholar
2021

Single-Shot Hyperspectral-Depth Imaging With Learned Diffractive Optics

ICCV 2021poster

Imaging depth and spectrum have been extensively studied in isolation from each other for decades. Recently, hyperspectral-depth (HS-D) imaging emerges to capture both information simultaneously by combining two different imaging systems; one for depth, the other for spectrum. While being accurate,…

Cited by 155PDFScholar
2020

Single-Shot Monocular RGB-D Imaging Using Uneven Double Refraction

CVPR 2020oral

Cameras that capture color and depth information have become an essential imaging modality for applications in robotics, autonomous driving, virtual, and augmented reality. Existing RGB-D cameras rely on multiple sensors or active illumination with specialized sensors. In this work, we propose a met…

Cited by 12PDFScholar
2018

Enhancing the Spatial Resolution of Stereo Images Using a Parallax Prior

CVPR 2018poster

We present a novel method that can enhance the spatial resolution of stereo images using a parallax prior. While traditional stereo imaging has focused on estimating depth from stereo images, our method utilizes stereo images to enhance spatial resolution instead of estimating disparity. The critica…

Cited by 162SourcePDFScholar