← Search

Satya Krishna Gorti

6 accepted papers

2025

MSc-SQL: Multi-Sample Critiquing Small Language Models For Text-To-SQL Translation

NAACL 2025long

Text-to-SQL generation enables non-experts to interact with databases via natural language. Recent advances rely on large closed-source models like GPT-4 that present challenges in accessibility, privacy, and latency. To address these issues, we focus on developing small, efficient, and open-source…

2024

Data-Efficient Multimodal Fusion on a Single GPU

CVPR 2024highlight

The goal of multimodal alignment is to learn a single latent space that is shared between multimodal inputs. The most powerful models in this space have been trained using massive datasets of paired inputs and large-scale computational resources making them prohibitively expensive to train in many p…

2023

TR0N: Translator Networks for 0-Shot Plug-and-Play Conditional Generation

ICML 2023poster

We propose TR0N, a highly general framework to turn pre-trained unconditional generative models, such as GANs and VAEs, into conditional models. The conditioning can be highly arbitrary, and requires only a pre-trained auxiliary model. For example, we show how to turn unconditional models into class…

2022

X-Pool: Cross-Modal Language-Video Attention for Text-Video Retrieval

CVPR 2022poster

In text-video retrieval, the objective is to learn a cross-modal similarity function between a text and a video that ranks relevant text-video pairs higher than irrelevant pairs. However, videos inherently express a much wider gamut of information than texts. Instead, texts often capture sub-regions…

Cited by 211PDFcodeScholar
2019

Guided Similarity Separation for Image Retrieval

NeurIPS 2019oral

Despite recent progress in computer vision, image retrieval remains a challenging open problem. Numerous variations such as view angle, lighting and occlusion make it difficult to design models that are both robust and efficient. Many leading methods traverse the nearest neighbor graph to exploit hi…