← Search

Si Shen

2 accepted papers

2026

Jointly Conditioned Diffusion Model for Multi-View Pose-Guided Person Image Synthesis

ICASSP 2026poster

Pose-guided human image generation is limited by incomplete textures from single reference views and the absence of explicit cross-view interaction. We present jointly conditioned diffusion model (JCDM), a jointly conditioned diffusion framework that exploits multi-view priors. The appearance prior…

Cited by 0SourcePDFScholar
2022

Increasing Visual Awareness in Multimodal Neural Machine Translation from an Information Theoretic Perspective

EMNLP 2022main

Multimodal machine translation (MMT) aims to improve translation quality by equipping the source sentence with its corresponding image. Despite the promising performance, MMT models still suffer the problem of input degradation: models focus more on textual information while visual information is ge…

Cited by 13SourcePDFScholar