← Search

Xinya Chen

6 accepted papers

2026

ScenDi: 3D-to-2D Scene Diffusion Cascades for Urban Generation

CVPR 2026

Recent advancements in 3D object generation using diffusion models have achieved remarkable success, but generating realistic 3D urban scenes remains challenging. Existing methods relying solely on 3D diffusion models tend to suffer a degradation in appearance details, while those utilizing only 2D

Cited by 0SourceScholar
2026

SemanticNVS: Improving Semantic Scene Understanding in Generative Novel View Synthesis

ICML 2026poster

We present SemanticNVS, a camera-conditioned multi-view diffusion model for novel view synthesis (NVS), which improves generation quality and consistency by integrating pre-trained semantic feature extractors. Existing NVS methods perform well for views near the input view, however, they tend to gen…

Cited by 0SourceScholar
2025

NormalCrafter: Learning Temporally Consistent Normals from Video Diffusion Priors

ICCV 2025poster

Surface normal estimation serves as a cornerstone for a spectrum of computer vision applications. While numerous efforts have been devoted to static image scenarios, ensuring temporal coherence in video-based normal estimation remains a formidable challenge. Instead of merely augmenting existing met…

2024

Learning 3D-aware GANs from Unposed Images with Template Feature Field

ECCV 2024oral

"Collecting accurate camera poses of training images has been shown to well serve the learning of 3D-aware generative adversarial networks (GANs) yet can be quite expensive in practice. This work targets learning 3D-aware GANs from unposed images, for which we propose to perform on-the-fly pose esti…

Cited by 1SourcePDFScholar
2023

VeRi3D: Generative Vertex-based Radiance Fields for 3D Controllable Human Image Synthesis

ICCV 2023poster

Unsupervised learning of 3D-aware generative adversarial networks has lately made much progress. Some recent work demonstrates promising results of learning human generative models using neural articulated radiance fields, yet their generalization ability and controllability lag behind parametric hu…

Cited by 9PDFScholar
2020

Adversarial Semantic Data Augmentation for Human Pose Estimation

ECCV 2020poster

Human pose estimation is the task of localizing body keypoints from still images. The state-of-the-art methods suffer from insufficient examples of challenging cases such as symmetric appearance, heavy occlusion and nearby person. To enlarge the amounts of challenging cases, previous methods augment…