← Search

Shreyas Kulkarni

3 accepted papers

2025

Understanding Figurative Meaning through Explainable Visual Entailment

NAACL 2025long

Large Vision-Language Models (VLMs) have demonstrated strong capabilities in tasks requiring a fine-grained understanding of literal meaning in images and text, such as visual question-answering or visual entailment. However, there has been little exploration of the capabilities of these models when…

2023

360FusionNeRF: Panoramic Neural Radiance Fields with Joint Guidance

IROS 2023poster

Based on the neural radiance fields (NeRF), we present a pipeline for generating novel views from a single 360° panoramic image. Prior research relied on the neighborhood interpolation capability of multi-layer perceptions to complete missing regions caused by occlusion. This resulted in artifacts i…

Cited by 24SourcecodeScholar
2023

Large Scale Generative Multimodal Attribute Extraction for E-commerce Attributes

ACL 2023industry

E-commerce websites (e.g. Amazon, Alibaba) have a plethora of structured and unstructured information (text and images) present on the product pages. Sellers often don’t label or mislabel values of the attributes (e.g. color, size etc.) for their products. Automatically identifying these attribute v…

Cited by 10SourcePDFScholar