← Search

Zejian Li

12 accepted papers

2026

Circular-DPO: Aligning Multi-Stage 3D Generative Models via Preference Feedback Loop

CVPR 2026

Multi-stage generative models have shown great promise in 3D content creation due to focused generation of structure or texture in different stages, but their outputs often fail to align with human preferences. The key bottleneck to apply alignment methods is the presence of non-differentiable opera

Cited by 0SourceScholar
2026

Diffusion Distillation with Direct Preference Optimization for Efficient 3D LiDAR Scene Completion

AAAI 2026technical

The slow sampling speed of diffusion models hinders their application in 3D LiDAR scene completion. To address this, we propose Distillation-DPO, a novel framework that accelerates sampling through score distillation while simultaneously enhancing generation quality via preference alignment. Disti

Cited by 0SourcePDFScholar
2026

Mean Flow Distillation: Robust and Stable Distillation for Flow Matching Models

ICML 2026poster

Flow Matching models have demonstrated strong performance across a wide range of generative tasks. However, their reliance on ODE-based iterative sampling incurs substantial computational overhead, which limits their applicability in real-time scenes. While distillation is a promising solution, exis…

Cited by 0SourceScholar
2026

When Diffusion Language Models Hesitate: Detecting and Correcting Visual Hallucinations via Confidence Fluctuation

ICML 2026poster

Multi-modal Diffusion Language Models (MDLMs) have emerged as a powerful alternative to autoregressive models in vision-language understanding, offering advantages in bidirectional context modeling and parallel decoding. However, existing MDLMs suffer from severe visual hallucinations due to the sta…

Cited by 0SourceScholar
2025

Distilling Diffusion Models to Efficient 3D LiDAR Scene Completion

ICCV 2025poster

Diffusion models have been applied to 3D LiDAR scene completion due to their strong training stability and high completion quality. However, the slow sampling speed limits the practical application of diffusion-based scene completion models since autonomous vehicles require an efficient perception o…

2025

Distribution Backtracking Builds A Faster Convergence Trajectory for Diffusion Distillation

ICLR 2025poster

Accelerating the sampling speed of diffusion models remains a significant challenge. Recent score distillation methods distill a heavy teacher model into a student generator to achieve one-step generation, which is optimized by calculating the difference between two score functions on the samples ge…

2024

Rapid 3D Model Generation with Intuitive 3D Input

CVPR 2024highlight

With the emergence of AR/VR 3D models are in tremendous demand. However conventional 3D modeling with Computer-Aided Design software requires much expertise and is difficult for novice users. We find that AR/VR devices in addition to serving as effective display mediums can offer a promising potenti…

Cited by 5SourcePDFScholar
2024

Reducing Spatial Fitting Error in Distillation of Denoising Diffusion Models

AAAI 2024technical

Denoising Diffusion models have exhibited remarkable capabilities in image generation. However, generating high-quality samples requires a large number of iterations. Knowledge distillation for diffusion models is an effective method to address this limitation with a shortened sampling process but c…

2023

Learning Object Consistency and Interaction in Image Generation from Scene Graphs

IJCAI 2023poster

This paper is concerned with synthesizing images conditioned on a scene graph (SG), a set of object nodes and their edges of interactive relations. We divide existing works into image-oriented and code-oriented methods. In our analysis, the image-oriented methods do not consider object interaction i…

2023

Preserving Structural Consistency in Arbitrary Artist and Artwork Style Transfer

AAAI 2023technical

Deep generative models are effective in style transfer. Previous methods learn one or several specific artist-style from a collection of artworks. These methods not only homogenize the artist-style of different artworks of the same artist but also lack generalization for the unseen artists. To solv…

Cited by 5SourcePDFScholar
2021

Image Synthesis From Layout With Locality-Aware Mask Adaption

ICCV 2021poster

This paper is concerned with synthesizing images conditioned on a layout (a set of bounding boxes with object categories). Existing works construct a layout-mask-image pipeline. Object masks are generated separately and mapped to bounding boxes to form a whole semantic segmentation mask (layout-to-m…

Cited by 80PDFcodeScholar