← Search

Dequan Wang*

2 accepted papers

2024

Dissecting Dissonance: Benchmarking Large Multimodal Models Against Self-Contradictory Instructions

ECCV 2024poster

"Large multimodal models (LMMs) excel in adhering to human instructions. However, self-contradictory instructions may arise due to the increasing trend of multimodal interaction and context length, which is challenging for language beginners and vulnerable populations. We introduce the Self-Contradi…

2024

Lost in Translation: Latent Concept Misalignment in Text-to-Image Diffusion Models

ECCV 2024poster

"Advancements in text-to-image diffusion models have broadened extensive downstream practical applications, but such models often encounter misalignment issues between text and image. Taking the generation of a combination of two disentangled concepts as an example, say given the prompt a tea cup of…