← Search

Tuhin Chakrabarty

20 accepted papers

2026

Death of the Novel(ty): Beyond N-Gram Novelty as a Metric for Textual Creativity

ICLR 2026poster

$N$-gram novelty is widely used to evaluate language models' ability to generate text outside of their training data. More recently, it has also been adopted as a metric for measuring textual creativity. However, theoretical work on creativity suggests that this approach may be inadequate, as it doe…

Cited by 0SourcecodeScholar
2025

Position: AI Safety should prioritize the Future of Work

ICML 2025oral

Current efforts in AI safety prioritize filtering harmful content, preventing manipulation of human behavior, and eliminating existential risks in cybersecurity or biosecurity. While pressing, this narrow focus overlooks critical human-centric considerations that shape the long-term trajectory of a…

Cited by 0SourcePDFScholar
2025

Understanding Figurative Meaning through Explainable Visual Entailment

NAACL 2025long

Large Vision-Language Models (VLMs) have demonstrated strong capabilities in tasks requiring a fine-grained understanding of literal meaning in images and text, such as visual question-answering or visual entailment. However, there has been little exploration of the capabilities of these models when…

2024

Connecting the Dots: Evaluating Abstract Reasoning Capabilities of LLMs Using the New York Times Connections Word Game

EMNLP 2024main

The New York Times Connections game has emerged as a popular and challenging pursuit for word puzzle enthusiasts. We collect438 Connections games to evaluate the performance of state-of-the-art large language models (LLMs) against expert and novice humanplayers. Our results show that even the best-p…

2024

Identifying Self-Disclosures of Use, Misuse and Addiction in Community-based Social Media Posts

NAACL 2024findings

In the last decade, the United States has lost more than 500,000 people from an overdose involving prescription and illicit opioids making it a national public health emergency (USDHHS, 2017). Medical practitioners require robust and timely tools that can effectively identify at-risk patients. Commu…

2023

I Spy a Metaphor: Large Language Models and Diffusion Models Co-Create Visual Metaphors

ACL 2023findings

Visual metaphors are powerful rhetorical devices used to persuade or communicate creative ideas through images. Similar to linguistic metaphors, they convey meaning implicitly through symbolism and juxtaposition of the symbols. We propose a new task of generating visual metaphors from linguistic met…

2023

Learning to Follow Object-Centric Image Editing Instructions Faithfully

EMNLP 2023long findings

Natural language instructions are a powerful interface for editing the outputs of text-to-image diffusion models. However, several challenges need to be addressed: 1) underspecification (the need to model the implicit meaning of instructions) 2) grounding (the need to localize where the edit has to…

Cited by 0SourcecodeScholar
2023

NORMSAGE: Multi-Lingual Multi-Cultural Norm Discovery from Conversations On-the-Fly

EMNLP 2023long main

Knowledge of norms is needed to understand and reason about acceptable behavior in human communication and interactions across sociocultural scenarios. Most computational research on norms has focused on a single culture, and manually built datasets, from non-conversational settings. We address thes…

Cited by 0SourcecodeScholar
2022

CONSISTENT: Open-Ended Question Generation From News Articles

EMNLP 2022finding

Recent work on question generation has largely focused on factoid questions such as who, what,where, when about basic facts. Generating open-ended why, how, what, etc. questions thatrequire long-form answers have proven more difficult. To facilitate the generation of openended questions, we propose…

2022

FLUTE: Figurative Language Understanding through Textual Explanations

EMNLP 2022main

Figurative language understanding has been recently framed as a recognizing textual entailment (RTE) task (a.k.a. natural language inference (NLI)). However, similar to classical RTE/NLI datasets they suffer from spurious correlations and annotation artifacts. To tackle this problem, work on NLI has…

2022

Help me write a poem: Instruction Tuning as a Vehicle for Collaborative Poetry Writing

EMNLP 2022main

Recent work in training large language models (LLMs) to follow natural language instructions has opened up exciting opportunities for natural language interface design. Building on the prior success of large language models in the realm of computer assisted creativity, in this work, we present CoPoe…

2022

Multitask Instruction-based Prompting for Fallacy Recognition

EMNLP 2022main

Fallacies are used as seemingly valid arguments to support a position and persuade the audience about its validity. Recognizing fallacies is an intrinsically difficult task both for humans and machines. Moreover, a big challenge for computational models lies in the fact that fallacies are formulated…

2021

COVID-Fact: Fact Extraction and Verification of Real-World Claims on COVID-19 Pandemic

ACL 2021long

We introduce a FEVER-like dataset COVID-Fact of 4,086 claims concerning the COVID-19 pandemic. The dataset contains claims, evidence for the claims, and contradictory claims refuted by the evidence. Unlike previous approaches, we automatically detect true claims and their source articles and then ge…

2021

DiSCoL: Toward Engaging Dialogue Systems through Conversational Line Guided Response Generation

NAACL 2021system demonstrations

Having engaging and informative conversations with users is the utmost goal for open-domain conversational systems. Recent advances in transformer-based language models and their applications to dialogue systems have succeeded to generate fluent and human-like responses. However, they still lack con…

Cited by 14SourcePDFScholar
2021

Don’t Go Far Off: An Empirical Study on Neural Poetry Translation

EMNLP 2021main

Despite constant improvements in machine translation quality, automatic poetry translation remains a challenging problem due to the lack of open-sourced parallel poetic corpora, and to the intrinsic complexities involved in preserving the semantics, style and figurative nature of poetry. We present…

2021

ENTRUST: Argument Reframing with Language Models and Entailment

NAACL 2021long

Framing involves the positive or negative presentation of an argument or issue depending on the audience and goal of the speaker. Differences in lexical framing, the focus of our work, can have large effects on peoples’ opinions and beliefs. To make progress towards reframing arguments for positive…

2021

Implicit Premise Generation with Discourse-aware Commonsense Knowledge Models

EMNLP 2021main

Enthymemes are defined as arguments where a premise or conclusion is left implicit. We tackle the task of generating the implicit premise in an enthymeme, which requires not only an understanding of the stated conclusion and premise but also additional inferences that could depend on commonsense kno…

2021

MERMAID: Metaphor Generation with Symbolism and Discriminative Decoding

NAACL 2021long

Generating metaphors is a challenging task as it requires a proper understanding of abstract concepts, making connections between unrelated concepts, and deviating from the literal meaning. In this paper, we aim to generate a metaphoric sentence given a literal expression by replacing relevant verbs…

2021

Metaphor Generation with Conceptual Mappings

ACL 2021long

Generating metaphors is a difficult task as it requires understanding nuanced relationships between abstract concepts. In this paper, we aim to generate a metaphoric sentence given a literal expression by replacing relevant verbs. Guided by conceptual metaphor theory, we propose to control the gener…