← Search

Xinmin Qiu

6 accepted papers

2026

Resolving Endpoint Underfitting in Diffusion Bridges via Noise Alignment

CVPR 2026

Diffusion bridge models offer a powerful framework for connecting two data distributions, such as in image restoration and translation. Many existing methods learn this bridge by mimicking the score-matching formulation of standard diffusion models. In this work, we find that this way leads to an an

Cited by 0SourcecodeScholar
2025

DreamHA: Towards High-Quality Human Animation with Image-to-Video Diffusion Models

ICASSP 2025accepted

Recent diffusion models have made significant advancements in generating lifelike videos from driving signals, including a reference character and a skeleton sequence. Nevertheless, these models often struggle with maintaining fidelity, as the generated results frequently deviate in character featur…

Cited by 0SourceScholar
2025

Feature out! Let Raw Image as Your Condition for Blind Face Restoration

ICML 2025poster

Blind face restoration (BFR), which involves converting low-quality (LQ) images into high-quality (HQ) images, remains challenging due to complex and unknown degradations. While previous diffusion-based methods utilize feature extractors from LQ images as guidance, using raw LQ images directly…

Cited by 0SourcePDFScholar
2025

MIRROR: Make Your Object-Level Multi-View Generation More Consistent with Training-Free Rectification

ICML 2025poster

Multi-view Diffusion has greatly advanced the development of 3D content creation by generating multiple images from distinct views, achieving remarkable photorealistic results. However, existing works are still vulnerable to inconsistent 3D geometric structures (commonly known as Janus Problem) and…

Cited by 0SourcePDFScholar
2025

StyO: Stylize Your Face in Only One-Shot

AAAI 2025technical

This paper focuses on face stylization with a single artistic target. Existing works for this task often fail to retain the source content while achieving geometry variation. Here, we present a novel StyO model, i.e., Stylize the face in only One-shot, to solve the above problem. In particular, StyO…

Cited by 9SourcePDFScholar
2024

BlazeBVD: Make Scale-Time Equalization Great Again for Blind Video Deflickering

ECCV 2024poster

"Developing blind video deflickering (BVD) algorithms to enhance video temporal consistency, is gaining importance amid the flourish of image processing and video generation. However, the intricate nature of video data complicates the training of deep learning methods, leading to high resource consu…

Cited by 0SourcePDFScholar