← Search

Stephen Rawls

3 accepted papers

2023

Translation-Enhanced Multilingual Text-to-Image Generation

ACL 2023long

Research on text-to-image generation (TTI) still predominantly focuses on the English language due to the lack of annotated image-caption data in other languages; in the long run, this might widen inequitable access to TTI technology. In this work, we thus investigate multilingual TTI (termed mTTI)…

2022

Multimodal Context Carryover

EMNLP 2022industry

Multi-modality support has become an integral part of creating a seamless user experience with modern voice assistants with smart displays. Users refer to images, video thumbnails, or the accompanying text descriptions on the screen through voice communication with AI powered devices. This raises th…

Cited by 3SourcePDFScholar