← Search

Hongxi Li

2 accepted papers

2026

Self-Prompting Diffusion Transformer for Open-Vocabulary Scene Text Edit via In-Context Learning

ICML 2026poster

Scene text editing aims to modify text in a target region of an image while preserving its background style and texture. Existing methods rely solely on image background information while neglecting the visual details of target regions, which discards stylistic features in the original text and esse…

Cited by 0SourceScholar
2025

Video Summarization Using Denoising Diffusion Probabilistic Model

AAAI 2025technical

Video summarization aims to eliminate visual redundancy while retaining key parts of video to construct concise and comprehensive synopses. Most existing methods use discriminative models to predict the importance scores of video frames. However, these methods are susceptible to annotation inconsist…

Cited by 0SourcePDFScholar