← Search

Vishwa Vinay

4 accepted papers

2022

CyCLIP: Cyclic Contrastive Language-Image Pretraining

NeurIPS 2022accept

Recent advances in contrastive representation learning over paired image-text data have led to models such as CLIP that achieve state-of-the-art performance for zero-shot classification and distributional robustness. Such models typically require joint reasoning in the image and text representation…

2022

Robustness of Fusion-based Multimodal Classifiers to Cross-Modal Content Dilutions

EMNLP 2022main

As multimodal learning finds applications in a wide variety of high-stakes societal tasks, investigating their robustness becomes important. Existing work has focused on understanding the robustness of vision-and-language models to imperceptible variations on benchmark tasks. In this work, we invest…

Cited by 8SourcePDFScholar
2022

VarScene: A Deep Generative Model for Realistic Scene Graph Synthesis

ICML 2022spotlight

Scene graphs are powerful abstractions that capture relationships between objects in images by modeling objects as nodes and relationships as edges. Generation of realistic synthetic scene graphs has applications like scene synthesis and data augmentation for supervised learning. Existing graph gene…

2021

Scene Graph Embeddings Using Relative Similarity Supervision

AAAI 2021technical

Scene graphs are a powerful structured representation of the underlying content of images, and embeddings derived from them have been shown to be useful in multiple downstream tasks. In this work, we employ a graph convolutional network to exploit structure in scene graphs and produce image embeddin…

Cited by 21SourcePDFScholar