← Search

Brandon Y. Feng

11 accepted papers

2026

First Frame Is the Place to Go for Video Content Customization

CVPR 2026

What role does the first frame play in video generation models? Traditionally, it's viewed as the spatial-temporal starting point of a video, merely a seed for subsequent animation. In this work, we reveal a fundamentally different perspective: video models implicitly treat the first frame as a conc

Cited by 0SourcecodeScholar
2025

IR3D-Bench: Evaluating Vision-Language Model Scene Understanding as Agentic Inverse Rendering

NeurIPS 2025poster

Vision-language models (VLMs) excel at descriptive tasks, but whether they truly understand scenes from visual observations remains uncertain. We introduce IR3D-Bench, a benchmark challenging VLMs to demonstrate understanding through active creation rather than passive recognition. Grounded in the a…

Cited by 0SourceScholar
2025

Parametric Shadow Control for Portrait Generation in Text-to-Image Diffusion Models

ICCV 2025poster

Text-to-image diffusion models excel at generating diverse portraits, but lack intuitive shadow control. Existing editing approaches, as post-processing, struggle to offer effective manipulation across diverse styles. Additionally, these methods either rely on expensive real-world light-stage data c…

Cited by 0SourcePDFScholar
2025

Repurposing Pre-trained Video Diffusion Models for Event-based Video Interpolation

CVPR 2025poster

Video Frame Interpolation aims to recover realistic missing frames between observed frames, generating a high-frame-rate video from a low-frame-rate video. However, without additional guidance, large motion between frames makes this problem ill-posed. Event-based Video Frame Interpolation (EVFI) add…

Cited by 4SourcePDFScholar
2024

Flash-Splat: 3D Reflection Removal with Flash Cues and Gaussian Splats

ECCV 2024poster

"We introduce a simple yet effective approach for separating transmitted and reflected light. Our key insight is that the powerful novel view synthesis capabilities provided by modern inverse rendering methods (e.g., 3D Gaussian splatting) allow one to perform flash/no-flash reflection separation us…

Cited by 8SourcePDFScholar
2024

Physics-Based Interaction with 3D Objects via Video Generation

ECCV 2024oral

"Realistic object interactions are crucial for creating immersive virtual experiences, yet synthesizing realistic 3D object dynamics in response to novel interactions remains a significant challenge. Unlike unconditional or text-conditioned dynamics generation, action-conditioned dynamics requires p…

2024

Temporally Consistent Atmospheric Turbulence Mitigation with Neural Representations

NeurIPS 2024poster

Atmospheric turbulence, caused by random fluctuations in the atmosphere's refractive index, introduces complex spatio-temporal distortions in imagery captured at long range. Video Atmospheric Turbulence Mitigation (ATM) aims to restore videos affected by these distortions. However, existing video AT…

2024

WaveMo: Learning Wavefront Modulations to See Through Scattering

CVPR 2024poster

Imaging through scattering media is a fundamental and pervasive challenge in fields ranging from medical diagnostics to astronomy. A promising strategy to overcome this challenge is wavefront modulation which induces measurement diversity during image acquisition. Despite its importance designing op…

2023

3D Motion Magnification: Visualizing Subtle Motions from Time-Varying Radiance Fields

ICCV 2023poster

Motion magnification helps us visualize subtle, imperceptible motion. However, prior methods only work for 2D videos captured with a fixed camera. We present a 3D motion magnification method that can magnify subtle motions from scenes captured by a moving camera, while supporting novel view renderin…

Cited by 7PDFScholar
2023

StegaNeRF: Embedding Invisible Information within Neural Radiance Fields

ICCV 2023poster

Recent advancements in neural rendering have paved the way for a future marked by the widespread distribution of visual data through the sharing of Neural Radiance Field (NeRF) model weights. However, while established techniques exist for embedding ownership or copyright information within conventi…

Cited by 66PDFcodeScholar
2022

PRIF: Primary Ray-Based Implicit Function

ECCV 2022poster

"We introduce a new implicit shape representation called Primary Ray-based Implicit Function (PRIF). In contrast to most existing approaches based on the signed distance function (SDF) which handles spatial locations, our representation operates on oriented rays. Specifically, PRIF is formulated to…

Cited by 56SourcePDFScholar