← Search

Khyathi Raghavi Chandu

3 accepted papers

2023

A Needle in a Haystack: An Analysis of High-Agreement Workers on MTurk for Summarization

ACL 2023long

To prevent the costly and inefficient use of resources on low-quality annotations, we want a method for creating a pool of dependable annotators who can effectively complete difficult tasks, such as evaluating automatic summarization. Thus, we investigate the recruitment of high-quality Amazon Mecha…

Cited by 11SourcePDFScholar
2022

Denoising Large-Scale Image Captioning from Alt-text Data Using Content Selection Models

COLING 2022main

Training large-scale image captioning (IC) models demands access to a rich and diverse set of training examples that are expensive to curate both in terms of time and man-power. Instead, alt-text based captions gathered from the web is a far cheaper alternative to scale with the downside of being no…

Cited by 2SourcePDFScholar
2021

Switch Point biased Self-Training: Re-purposing Pretrained Models for Code-Switching

EMNLP 2021finding

Code-switching (CS), a ubiquitous phenomenon due to the ease of communication it offers in multilingual communities still remains an understudied problem in language processing. The primary reasons behind this are: (1) minimal efforts in leveraging large pretrained multilingual models, and (2) the l…

Cited by 5SourcePDFScholar