← Search

Niyati Chhaya

10 accepted papers

2024

IndicIRSuite: Multilingual Dataset and Neural Information Models for Indian Languages

ACL 2024short

In this paper, we introduce Neural Information Retrieval resources for 11 widely spoken Indian Languages (Assamese, Bengali, Gujarati, Hindi, Kannada, Malayalam, Marathi, Oriya, Punjabi, Tamil, and Telugu) from two major Indian language families (Indo-Aryan and Dravidian). These resources include (a…

2023

Open-World Factually Consistent Question Generation

ACL 2023findings

Question generation methods based on pre-trained language models often suffer from factual inconsistencies and incorrect entities and are not answerable from the input paragraph. Domain shift – where the test data is from a different domain than the training data - further exacerbates the problem of…

Cited by 4SourcePDFScholar
2023

Sketch Recognition via Part-based Hierarchical Analogical Learning

IJCAI 2023poster

Sketch recognition has been studied for decades, but it is far from solved. Drawing styles are highly variable across people and adapting to idiosyncratic visual expressions requires data-efficient learning. Explainability also matters, so that users can see why a system got confused about something…

Cited by 4SourcePDFScholar
2023

“Let’s not Quote out of Context”: Unified Vision-Language Pretraining for Context Assisted Image Captioning

ACL 2023industry

Well-formed context aware image captions and tags in enterprise content such as marketing material are critical to ensure their brand presence and content recall. Manual creation and updates to ensure the same is non trivial given the scale and the tedium towards this task. We propose a new unified…

Cited by 8SourcePDFScholar
2022

CaM-Gen: Causally Aware Metric-Guided Text Generation

ACL 2022findings

Content is created for a well-defined purpose, often described by a metric or signal represented in the form of structured information. The relationship between the goal (metrics) of target content and the content itself is non-trivial. While large-scale language models show promising text generatio…

2022

Offer a Different Perspective: Modeling the Belief Alignment of Arguments in Multi-party Debates

EMNLP 2022main

In contexts where debate and deliberation are the norm, the participants are regularly presented with new information that conflicts with their original beliefs. When required to update their beliefs (belief alignment), they may choose arguments that align with their worldview (confirmation bias). W…

2021

AUTOSUMM: Automatic Model Creation for Text Summarization

EMNLP 2021main

Recent efforts to develop deep learning models for text generation tasks such as extractive and abstractive summarization have resulted in state-of-the-art performances on various datasets. However, obtaining the best model configuration for a given dataset requires an extensive knowledge of deep le…

2021

Counterfactuals to Control Latent Disentangled Text Representations for Style Transfer

ACL 2021short

Disentanglement of latent representations into content and style spaces has been a commonly employed method for unsupervised text style transfer. These techniques aim to learn the disentangled representations and tweak them to modify the style of a sentence. In this paper, we propose a counterfactua…

2021

WikiTalkEdit: A Dataset for modeling Editors’ behaviors on Wikipedia

NAACL 2021long

This study introduces and analyzes WikiTalkEdit, a dataset of conversations and edit histories from Wikipedia, for research in online cooperation and conversation modeling. The dataset comprises dialog triplets from the Wikipedia Talk pages, and editing actions on the corresponding articles being di…

2020

Semi-supervised Multi-task Learning for Multi-label Fine-grained Sexism Classification

COLING 2020main

Sexism, a form of oppression based on one’s sex, manifests itself in numerous ways and causes enormous suffering. In view of the growing number of experiences of sexism reported online, categorizing these recollections automatically can assist the fight against sexism, as it can facilitate effective…