← Search

Ruidong Chen

5 accepted papers

2026

Layer-wise Instance Binding for Regional and Occlusion Control in Text-to-Image Diffusion Transformers

CVPR 2026

Region-instructed layout control in text-to-image generation is highly practical, yet existing methods suffer from limitations: (i) training-based approaches inherit data bias and often degrade image quality, and (ii) current techniques struggle with occlusion order, limiting real-world usability. T

Cited by 0SourcecodeScholar
2026

T2I-RiskyPrompt: A Benchmark for Safety Evaluation, Attack, and Defense on Text-to-Image Model

AAAI 2026technical

Using risky text prompts, such as pornography and violent prompts, to test the safety of text-to-image (T2I) models is a critical task. However, existing risky prompt datasets are limited in three key areas: 1) limited risky categories, 2) coarse-grained annotation, and 3) low effectiveness. To addr

Cited by 0SourcePDFScholar
2025

Aesthetic Perception Prompting for Interpretable Image Aesthetics Assessment with MLLMs

ICASSP 2025accepted

Image Aesthetic Assessment (IAA) aims to rate the aesthetic quality of images and has many practical applications. However, existing methods typically rely on limited annotated data for training, leading to two key issues: 1) score-only predictions lack interpretability, making it hard for users to…

Cited by 0SourceScholar
2025

TRCE: Towards Reliable Malicious Concept Erasure in Text-to-Image Diffusion Models

ICCV 2025poster

Recent advances in text-to-image diffusion models enable photorealistic image generation, but they also risk producing malicious content, such as NSFW images. To mitigate risk, concept erasure methods are studied to facilitate the model to unlearn specific concepts. However, current studies struggle…

2024

AnyScene: Customized Image Synthesis with Composited Foreground

CVPR 2024poster

Recent advancements in text-to-image technology have significantly advanced the field of image customization. Among various applications the task of customizing diverse scenes for user-specified composited elements holds great application value but has not been extensively explored. Addressing this…

Cited by 1SourcePDFScholar