← Search

Vijit Malik

4 accepted papers

2024

CorrSynth - A Correlated Sampling Method for Diverse Dataset Generation from LLMs

EMNLP 2024main

Large language models (LLMs) have demonstrated remarkable performance in diverse tasks using zero-shot and few-shot prompting. Even though their capabilities of data synthesis have been studied well in recent years, the generated data suffers from a lack of diversity, less adherence to the prompt, a…

Cited by 1SourcePDFScholar
2024

PEARL: Preference Extraction with Exemplar Augmentation and Retrieval with LLM Agents

EMNLP 2024industry

Identifying preferences of customers in their shopping journey is a pivotal aspect in providing product recommendations. The task becomes increasingly challenging when there is a multi-turn conversation between the user and a shopping assistant chatbot. In this paper, we tackle a novel and complex p…

2022

Socially Aware Bias Measurements for Hindi Language Representations

NAACL 2022long

Language representations are an efficient tool used across NLP, but they are strife with encoded societal biases. These biases are studied extensively, but with a primary focus on English language representations and biases common in the context of Western society. In this work, we investigate the b…

2021

ILDC for CJPE: Indian Legal Documents Corpus for Court Judgment Prediction and Explanation

ACL 2021long

An automated system that could assist a judge in predicting the outcome of a case would help expedite the judicial process. For such a system to be practically useful, predictions by the system should be explainable. To promote research in developing such a system, we introduce ILDC (Indian Legal Do…