← Search

Shanyuan Liu

6 accepted papers

2026

NAMI: Efficient Image Generation via Bridged Progressive Rectified Flow Transformers

CVPR 2026

Flow-based Transformer models have achieved state-of-the-art image generation performance, but often suffer from high inference latency and computational cost due to their large parameter sizes. To improve inference efficiency without compromising quality, we propose Bridged Progressive Rectified Fl

Cited by 0SourceScholar
2026

RefTon: Reference person shot assist virtual Try-on

CVPR 2026

We introduce RefTon, a flux-based person-to-person virtual try-on framework that enhances garment realism through unpaired visual references. Unlike conventional approaches that rely on complex auxiliary inputs such as body parsing and warped mask or require finely designed extract branches to proce

Cited by 0SourcecodeScholar
2026

RevealLayer: Disentangling Hidden and Visible Layers via Occlusion-Aware Image Decomposition

ICML 2026poster

Recent diffusion-based approaches have made substantial progress in image layer decomposition. However, accurately decomposing complex natural images remains challenging due to difficulties in occlusion completion, robust layer disentanglement, and precise foreground boundaries. Moreover, the scarci…

Cited by 0SourceScholar
2025

Bridge Diffusion Model: Bridge Chinese Text-to-Image Diffusion Model with English Communities

AAAI 2025technical

Text-to-Image generation (TTI) technologies are advancing rapidly, especially in the English language communities. However, apart from the user input language barrier problem, English-native TTI models inherently carry biases from their English world centric training data, which creates a dilemma fo…

2025

PlanGen: Towards Unified Layout Planning and Image Generation in Auto-Regressive Vision Language Models

ICCV 2025poster

In this paper, we propose a unified layout planning and image generation model, PlanGen, which can pre-plan spatial layout conditions before generating images as shown in Figure 1. Unlike previous diffusion-based models that treat layout planning and layout-to-image as two separate models, PlanGen j…

Cited by 0SourcePDFScholar
2024

HiCo: Hierarchical Controllable Diffusion Model for Layout-to-image Generation

NeurIPS 2024poster

The task of layout-to-image generation involves synthesizing images based on the captions of objects and their spatial positions. Existing methods still struggle in complex layout generation, where common bad cases include object missing, inconsistent lighting, conflicting view angles, etc. To effec…