← Search

Shen Sang

7 accepted papers

2026

Video-As-Prompt: Unified Semantic Control for Video Generation

ICLR 2026poster

Unified, generalizable semantic control in video generation remains a critical open challenge. Existing methods either introduce artifacts by enforcing inappropriate pixel-wise priors from structure-based controls, or rely on non-generalizable, condition-specific finetuning or task-specific architec…

Cited by 0SourcecodeScholar
2025

COAP: Memory-Efficient Training with Correlation-Aware Gradient Projection

CVPR 2025poster

Training large-scale neural networks in vision, and multimodal domains demands substantial memory resources, primarily due to the storage of optimizer states. While LoRA, a popular parameter-efficient method, reduces memory usage, it often suffers from suboptimal performance due to the constraints o…

Cited by 3SourcePDFScholar
2025

ID-Patch: Robust ID Association for Group Photo Personalization

CVPR 2025poster

The ability to synthesize personalized group photos and specify the positions of each identity offers immense creative potential. While such imagery can be visually appealing, it presents significant challenges for existing technologies. A persistent issue is identity (ID) leakage, where injected fa…

2023

ActorsNeRF: Animatable Few-shot Human Rendering with Generalizable NeRFs

ICCV 2023poster

While NeRF-based human representations have shown impressive novel view synthesis results, most methods still rely on a large number of images / views for training. In this work, we propose a novel animatable NeRF called ActorsNeRF. It is first pre-trained on diverse human subjects, and then adapted…

Cited by 25PDFScholar
2021

OpenRooms: An Open Framework for Photorealistic Indoor Scene Datasets

CVPR 2021poster

We propose a novel framework for creating large-scale photorealistic datasets of indoor scenes, with ground truth geometry, material, lighting and semantics. Our goal is to make the dataset creation process widely accessible, allowing researchers to transform scans into datasets with highquality gro…

Cited by 93PDFScholar