← Search

Ashutosh Modi

23 accepted papers

2025

Beyond Components: Singular Vector-Based Interpretability of Transformer Circuits

NeurIPS 2025poster

Transformer-based language models exhibit complex behavior, but their internal computations remain poorly understood. Most mechanistic interpretability approaches treat components, such as attention heads and MLPs, as atomic units, ignoring potential functional substructure. We propose a finer-grain…

Cited by 0SourceScholar
2025

CoMuMDR: Code-mixed Multi-modal Multi-domain corpus for Discourse paRsing in conversations

ACL 2025finding

Discourse parsing is an important task useful for NLU applications such as summarization, machine comprehension, and emotion recognition. The current discourse parsing datasets based on conversations consists of written English dialogues restricted to a single domain. In this resource paper, we intr…

2025

EtiCor++: Towards Understanding Etiquettical Bias in LLMs

ACL 2025finding

In recent years, researchers have started analyzing the cultural sensitivity of LLMs. In this respect, Etiquettes have been an active area of research. Etiquettes are region-specific and are an essential part of the culture of a region; hence, it is imperative to make LLMs sensitive to etiquettes. H…

2025

IL-PCSR: Legal Corpus for Prior Case and Statute Retrieval

EMNLP 2025

Identifying/retrieving relevant statutes and prior cases/precedents for a given legal situation are common tasks exercised by law practitioners. Researchers till date have addressed the two tasks independently, thus developing completely different datasets and models for each task; however, both ret

2025

PoseStitch-SLT: Linguistically Inspired Pose-Stitching for End-to-End Sign Language Translation

EMNLP 2025

Sign language translation remains a challenging task due to the scarcity of large-scale, sentence-aligned datasets. Prior arts have focused on various feature extraction and architectural changes to support neural machine translation for sign languages. We propose PoseStitch-SLT, a novel pre-trainin

Cited by 0SourcePDFScholar
2024

BookSQL: A Large Scale Text-to-SQL Dataset for Accounting Domain

NAACL 2024long

Several large-scale datasets (e.g., WikiSQL, Spider) for developing natural language interfaces to databases have recently been proposed. These datasets cover a wide breadth of domains but fall short on some essential domains, such as finance and accounting. Given that accounting databases are used…

2024

IL-TUR: Benchmark for Indian Legal Text Understanding and Reasoning

ACL 2024long

Legal systems worldwide are inundated with exponential growth in cases and documents. There is an imminent need to develop NLP and ML techniques for automatically processing and understanding legal documents to streamline the legal system. However, evaluating and comparing various NLP models designe…

2024

Towards Measuring and Modeling “Culture” in LLMs: A Survey

EMNLP 2024main

We present a survey of more than 90 recent papers that aim to study cultural representation and inclusion in large language models (LLMs). We observe that none of the studies explicitly define “culture, which is a complex, multifaceted concept; instead, they probe the models on some specially design…

2024

Towards Robust Evaluation of Unlearning in LLMs via Data Transformations

EMNLP 2024finding

Large Language Models (LLMs) have shown to be a great success in a wide range of applications ranging from regular NLP-based use cases to AI agents. LLMs have been trained on a vast corpus of texts from various sources; despite the best efforts during the data pre-processing stage while training the…

2024

iSign: A Benchmark for Indian Sign Language Processing

ACL 2024findings

Indian Sign Language has limited resources for developing machine learning and data-driven approaches for automated language processing. Though text/audio-based language processing techniques have shown colossal research interest and tremendous improvements in the last few years, Sign Languages stil…

2023

ScriptWorld: Text Based Environment for Learning Procedural Knowledge

IJCAI 2023poster

Text-based games provide a framework for developing natural language understanding and commonsense knowledge about the world in reinforcement learning based agents. Existing text-based environments often rely on fictional situations and characters to create a gaming framework and are far from real-w…

2023

U-CREAT: Unsupervised Case Retrieval using Events extrAcTion

ACL 2023long

The task of Prior Case Retrieval (PCR) in the legal domain is about automatically citing relevant (based on facts and precedence) prior legal cases in a given query case. To further promote research in PCR, in this paper, we propose a new large benchmark (in English) for the PCR task: IL-PCR (Indian…

2022

CISLR: Corpus for Indian Sign Language Recognition

EMNLP 2022main

Indian Sign Language, though used by a diverse community, still lacks well-annotated resources for developing systems that would enable sign language processing. In recent years researchers have actively worked for sign languages like American Sign Languages, however, Indian Sign language is still f…

Cited by 12SourcePDFScholar
2022

COGMEN: COntextualized GNN based Multimodal Emotion recognitioN

NAACL 2022long

Emotions are an inherent part of human interactions, and consequently, it is imperative to develop AI systems that understand and recognize human emotions. During a conversation involving various people, a person’s emotions are influenced by the other speaker’s utterances and their own emotional sta…

2022

HLDC: Hindi Legal Documents Corpus

ACL 2022findings

Many populous countries including India are burdened with a considerable backlog of legal cases. Development of automated systems that could process legal documents and augment legal practitioners can mitigate this. However, there is a dearth of high-quality corpora that is needed to develop such da…

2021

ILDC for CJPE: Indian Legal Documents Corpus for Court Judgment Prediction and Explanation

ACL 2021long

An automated system that could assist a judge in predicting the outcome of a case would help expedite the judicial process. For such a system to be practically useful, predictions by the system should be explainable. To promote research in developing such a system, we introduce ILDC (Indian Legal Do…

2020

Adapting a Language Model for Controlled Affective Text Generation

COLING 2020main

Human use language not just to convey information but also to express their inner feelings and mental states. In this work, we adapt the state-of-the-art language generation models to generate affective (emotional) text. We posit a model capable of generating affect-driven and topic focused sentence…