← Search

Ruoyu Zhao

4 accepted papers

2025

Diff-2-in-1: Bridging Generation and Dense Perception with Diffusion Models

ICLR 2025poster

Beyond high-fidelity image synthesis, diffusion models have recently exhibited promising results in dense visual perception tasks. However, most existing work treats diffusion models as a standalone component for perception tasks, employing them either solely for off-the-shelf data augmentation or a…

2025

Textualize Visual Prompt for Image Editing via Diffusion Bridge

AAAI 2025technical

Visual prompt, a pair of before-and-after edited images, can convey indescribable imagery transformations and prosper in image editing. However, current visual prompt methods rely on a pretrained text-guided image-to-image generative model that requires a triplet of text, before, and after images fo…

Cited by 0SourcePDFScholar
2024

FreeDiff: Progressive Frequency Truncation for Image Editing with Diffusion Models

ECCV 2024poster

"Precise image editing with text-to-image models has attracted increasing interest due to their remarkable generative capabilities and user-friendly nature. However, such attempts face the pivotal challenge of misalignment between the intended precise editing target regions and the broader area impa…

2021

Hierarchical Multi-label Text Classification with Horizontal and Vertical Category Correlations

EMNLP 2021main

Hierarchical multi-label text classification (HMTC) deals with the challenging task where an instance can be assigned to multiple hierarchically structured categories at the same time. The majority of prior studies either focus on reducing the HMTC task into a flat multi-label problem ignoring the v…

Cited by 31SourcePDFScholar