← Search

Nishtha Madaan

3 accepted papers

2024

LLMGuard: Guarding against Unsafe LLM Behavior

AAAI 2024technical

Although the rise of Large Language Models (LLMs) in enterprise settings brings new opportunities and capabilities, it also brings challenges, such as the risk of generating inappropriate, biased, or misleading content that violates regulations and can have legal concerns. To alleviate this, we pres…

Cited by 11SourcePDFScholar
2023

DetAIL: A Tool to Automatically Detect and Analyze Drift in Language

AAAI 2023technical

Machine learning and deep learning-based decision making has become part of today's software. The goal of this work is to ensure that machine learning and deep learning-based systems are as trusted as traditional software. Traditional software is made dependable by following rigorous practice like s…

Cited by 6SourcePDFScholar
2021

Generate Your Counterfactuals: Towards Controlled Counterfactual Generation for Text

AAAI 2021technical

Machine Learning has seen tremendous growth recently, which has led to a larger adaptation of ML systems for educational assessments, credit risk, healthcare, employment, criminal justice, to name a few. The trustworthiness of ML and NLP systems is a crucial aspect and requires a guarantee that the…

Cited by 111SourcePDFScholar