← Search

Minye Wu

15 accepted papers

2025

BG-Triangle: Bezier Gaussian Triangle for 3D Vectorization and Rendering

CVPR 2025poster

Differentiable rendering enables efficient optimization by allowing gradients to be computed through the rendering process, facilitating 3D reconstruction, inverse rendering and neural scene representation learning. To ensure differentiability, existing solutions approximate or re-formulate traditio…

Cited by 1SourcePDFScholar
2024

Navigating the Nuances: A Fine-grained Evaluation of Vision-Language Navigation

EMNLP 2024finding

This study presents a novel evaluation framework for the Vision-Language Navigation (VLN) task. It aims to diagnose current models for various instruction categories at a finer-grained level. The framework is structured around the context-free grammar (CFG) of the task. The CFG serves as the basis f…

2024

TeTriRF: Temporal Tri-Plane Radiance Fields for Efficient Free-Viewpoint Video

CVPR 2024poster

Neural Radiance Fields (NeRF) revolutionize the realm of visual media by providing photorealistic Free-Viewpoint Video (FVV) experiences offering viewers unparalleled immersion and interactivity. However the technology's significant storage requirements and the computational complexity involved in g…

Cited by 18SourcePDFScholar
2024

VideoRF: Rendering Dynamic Radiance Fields as 2D Feature Video Streams

CVPR 2024poster

Neural Radiance Fields (NeRFs) excel in photorealistically rendering static scenes. However rendering dynamic long-duration radiance fields on ubiquitous devices remains challenging due to data storage and computational constraints. In this paper we introduce VideoRF the first approach to enable rea…

2023

Neural Residual Radiance Fields for Streamably Free-Viewpoint Videos

CVPR 2023poster

The success of the Neural Radiance Fields (NeRFs) for modeling and free-view rendering static objects has inspired numerous attempts on dynamic scenes. Current techniques that utilize neural rendering for facilitating free-view videos (FVVs) are restricted to either offline rendering or are capable…

Cited by 66SourcePDFScholar
2022

Fourier PlenOctrees for Dynamic Radiance Field Rendering in Real-Time

CVPR 2022oral

Implicit neural representations such as Neural Radiance Field (NeRF) have focused mainly on modeling static objects captured under multi-view settings where real-time rendering can be achieved with smart data structures, e.g., PlenOctree. In this paper, we present a novel Fourier PlenOctree (FPO) te…

Cited by 178PDFScholar
2022

NeuralHOFusion: Neural Volumetric Rendering Under Human-Object Interactions

CVPR 2022poster

4D modeling of human-object interactions is critical for numerous applications. However, efficient volumetric capture and rendering of complex interaction scenarios, especially from sparse inputs, remain challenging. In this paper, we propose NeuralHOFusion, a neural approach for volumetric human-ob…

Cited by 50PDFScholar
2021

ChallenCap: Monocular 3D Capture of Challenging Human Performances Using Multi-Modal References

CVPR 2021poster

Capturing challenging human motions is critical for numerous applications, but it suffers from complex motion patterns and severe self-occlusion under the monocular setting. In this paper, we propose ChallenCap --- a template-based approach to capture challenging 3D human motions using a single RGB…

Cited by 28PDFScholar
2021

Few-shot Neural Human Performance Rendering from Sparse RGBD Videos

IJCAI 2021poster

Recent neural rendering approaches for human activities achieve remarkable view synthesis results, but still rely on dense input views or dense training with all the capture frames, leading to deployment difficulty and inefficient training overload. However, existing advances will be ill-posed if th…

Cited by 17SourcePDFScholar
2021

GNeRF: GAN-Based Neural Radiance Field Without Posed Camera

ICCV 2021poster

We introduce GNeRF, a framework to marry Generative Adversarial Networks (GAN) with Neural Radiance Field (NeRF) reconstruction for the complex scenarios with unknown and even randomly initialized camera poses. Recent NeRF-based advances have gained popularity for remarkable realistic novel view syn…

Cited by 223PDFcodeScholar
2021

Neural Video Portrait Relighting in Real-Time via Consistency Modeling

ICCV 2021poster

Video portraits relighting is critical in user-facing human photography, especially for immersive VR/AR experience. Recent advances still fail to recover consistent relit result under dynamic illuminations from monocular RGB stream, suffering from the lack of video consistency supervision. In this p…

Cited by 47PDFcodeScholar
2021

NeuralHumanFVV: Real-Time Neural Volumetric Human Performance Rendering Using RGB Cameras

CVPR 2021poster

4D reconstruction and rendering of human activities is critical for immersive VR/AR experience. Recent advances still fail to recover fine geometry and texture results with the level of detail present in the input images from sparse multi-view RGB cameras. In this paper, we propose NeuralHumanFVV, a…

Cited by 50PDFScholar
2021

PIANO: A Parametric Hand Bone Model from Magnetic Resonance Imaging

IJCAI 2021poster

Hand modeling is critical for immersive VR/AR, action understanding, or human healthcare. Existing parametric models account only for hand shape, pose, or texture, without modeling the anatomical attributes like bone, which is essential for realistic hand biomechanics analysis. In this paper, we pre…