← Search

Shaocheng Shen

1 accepted papers

2026

Agentic Retoucher for Text-To-Image Generation

CVPR 2026

Text-to-image (T2I) diffusion models such as SDXL and FLUX have achieved impressive photorealism, yet small-scale distortions remain pervasive in limbs, face, text and so on. Existing refinement approaches either perform costly iterative re-generation or rely on vision-language models (VLMs) with we

Cited by 3SourcecodeScholar