← Search

Anay Mehrotra

11 accepted papers

2026

Language Generation with Feedback: Queries and Mistakes

ICML 2026poster

We investigate language generation in the limit (Kleinberg & Mullainathan, 2024; Li et al., 2025) in variants where the generator receives some feedback based on its “actions.” We study two such variants. In the first, which is inspired by Littlestone’s model of online learning, the generator observ…

Cited by 0SourceScholar
2026

Linear Regression with Unknown Truncation Beyond Gaussian Features

ICML 2026poster

In truncated linear regression, samples $(x,y)$ are shown only when the outcome $y$ falls inside a certain survival set $S^\star$ and the goal is to estimate the unknown $d$-dimensional regressor $w^\star$. This problem has a long history of study in Statistics and Machine Learning going back to the…

Cited by 0SourceScholar
2026

Mean Estimation from Coarse Data: Characterizations and Efficient Algorithms

ICLR 2026poster

Coarse data arise when learners observe only partial information about samples; namely, a set containing the sample rather than its exact value. This occurs naturally through measurement rounding, sensor limitations, and lag in economic systems. We study Gaussian mean estimation from coarse data, wh…

Cited by 0SourceScholar
2024

Fair Classification with Partial Feedback: An Exploration-Based Data Collection Approach

ICML 2024poster

In many predictive contexts (e.g., credit lending), true outcomes are only observed for samples that were positively classified in the past. These past observations, in turn, form training datasets for classifiers that make future predictions. However, such training datasets lack information about t…

2024

Tree of Attacks: Jailbreaking Black-Box LLMs Automatically

NeurIPS 2024poster

While Large Language Models (LLMs) display versatile functionality, they continue to generate harmful, biased, and toxic content, as demonstrated by the prevalence of human-designed *jailbreaks*. In this work, we present *Tree of Attacks with Pruning* (TAP), an automated method for generating jailb…

2023

Bias in Evaluation Processes: An Optimization-Based Model

NeurIPS 2023poster

Biases with respect to socially-salient attributes of individuals have been well documented in evaluation processes used in settings such as admissions and hiring. We view such an evaluation process as a transformation of a distribution of the true utility of an individual for a task to an observed…

2023

Subset Selection Based On Multiple Rankings in the Presence of Bias: Effectiveness of Fairness Constraints for Multiwinner Voting Score Functions

ICML 2023poster

We consider the problem of subset selection where one is given multiple rankings of items and the goal is to select the highest "quality" subset. Score functions from the multiwinner voting literature have been used to aggregate rankings into quality scores for subsets. We study this setting of subs…