← Search

Alchan Hwang

1 accepted papers

2025

Exploring Multimodal Diffusion Transformers for Enhanced Prompt-based Image Editing

ICCV 2025poster

Transformer-based diffusion models have recently superseded traditional U-Net architectures, with multimodal diffusion transformers (MM-DiT) emerging as the dominant approach in state-of-the-art models like Stable Diffusion 3 and Flux.1. Previous approaches have relied on unidirectional cross-attent…

Cited by 0SourcePDFScholar