← Search

Hualian Sheng

8 accepted papers

2026

AnyID: Ultra-Fidelity Universal Identity-Preserving Video Generation from Any Visual References

CVPR 2026

Identity-preserving video generation offers powerful tools for creative expression, allowing users to customize videos featuring their beloved characters. However, prevailing methods are typically designed and optimized for a single identity reference. This underlying assumption restricts creative f

Cited by 0SourcecodeScholar
2026

EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer

ICLR 2026poster

Video generation models have advanced significantly, yet they still struggle to synthesize complex human movements due to the high degrees of freedom in human articulation. This limitation stems from the intrinsic constraints of pixel-only training objectives, which inherently bias models toward app…

Cited by 0SourceScholar
2025

EchoShot: Multi-Shot Portrait Video Generation

NeurIPS 2025poster

Video diffusion models substantially boost the productivity of artistic workflows with high-quality portrait video generative capacity. However, prevailing pipelines are primarily constrained to single-shot creation, while real-world applications urge for multiple shots with identity consistency and…

Cited by 0SourcecodeScholar
2025

PerLDiff: Controllable Street View Synthesis Using Perspective-Layout Diffusion Model

ICCV 2025poster

Controllable generation is considered a potentially vital approach to address the challenge of annotating 3D data, and the precision of such controllable generation becomes particularly imperative in the context of data production for autonomous driving. Existing methods focus on the integration of…

2024

RoScenes: A Large-scale Multi-view 3D Dataset for Roadside Perception

ECCV 2024poster

"We introduce RoScenes, the largest multi-view roadside perception dataset, which aims to shed light on the development of vision-centric Bird’s Eye View (BEV) approaches for more challenging traffic scenes. The highlights of RoScenes include significantly large perception area, full scene coverage…

2022

Balanced and Hierarchical Relation Learning for One-Shot Object Detection

CVPR 2022poster

Instance-level feature matching is significantly important to the success of modern one-shot object detectors. Recently, the methods based on the metric-learning paradigm have achieved an impressive process. Most of these works only measure the relations between query and target objects on a single…

Cited by 31PDFcodeScholar
2022

Rethinking IoU-Based Optimization for Single-Stage 3D Object Detection

ECCV 2022poster

"Since Intersection-over-Union (IoU) based optimization maintains the consistency of the final IoU prediction metric and losses, it has been widely used in both regression and classification branches of single-stage 2D object detectors. Recently, several 3D object detection methods adopt IoU-based o…

2021

Improving 3D Object Detection With Channel-Wise Transformer

ICCV 2021poster

Though 3D object detection from point clouds has achieved rapid progress in recent years, the lack of flexible and high-performance proposal refinement remains a great hurdle for existing state-of-the-art two-stage detectors. Previous works on refining 3D proposals have relied on human-designed comp…

Cited by 298PDFcodeScholar