← Search

Hareesh Ravi

3 accepted papers

2022

Cross-Modal Coherence for Text-to-Image Retrieval

AAAI 2022technical

Common image-text joint understanding techniques presume that images and the associated text can universally be characterized by a single implicit model. However, co-occurring images and text can be related in qualitatively different ways, and explicitly modeling it could improve the performance of…

2021

AESOP: Abstract Encoding of Stories, Objects, and Pictures

ICCV 2021poster

Visual storytelling and story comprehension are uniquely human skills that play a central role in how we learn about and experience the world. Despite remarkable progress in recent years in synthesis of visual and textual content in isolation and learning effective joint visual-linguistic representa…

Cited by 19PDFcodeScholar
2018

Show Me a Story: Towards Coherent Neural Story Illustration

CVPR 2018poster

We propose an end-to-end network for the visual illustration of a sequence of sentences forming a story. At the core of our model is the ability to model the inter-related nature of the sentences within a story, as well as the ability to learn coherence to support reference resolution. The framework…