← Search

Yaman Kumar Singla

8 accepted papers

2026

ALPHA: Action-Based Learning for Pluralistic Human Alignment in Large Language Models

AAAI 2026technical

Large language models are widely used, but aligning them with societal values remains challenging. Current approaches often rely on human annotations, which are hard to scale, or synthetic data produced by models that may themselves be misaligned, making it difficult to capture genuine public opinio

Cited by 0SourcePDFScholar
2026

Social Agents: Collective Intelligence Improves LLM Predictions

ICLR 2026poster

In human society, collective decision making has often outperformed the judgment of individuals. Classic examples range from estimating livestock weights to predicting elections and financial markets, where averaging many independent guesses often yields results more accurate than experts. These suc…

Cited by 0SourceScholar
2025

Measuring And Improving Engagement of Text-to-Image Generation Models

ICLR 2025poster

Recent advances in text-to-image generation have achieved impressive aesthetic quality, making these models usable for both personal and commercial purposes. However, in the fields of marketing and advertising, images are often created to be more engaging, as reflected in user behaviors such as incr…

2025

Measuring And Improving Persuasiveness Of Large Language Models

ICLR 2025poster

Large Language Models (LLMs) are increasingly being used in workflows involving generating content to be consumed by humans (*e.g.,* marketing) and also in directly interacting with humans (*e.g.,* through chatbots). The development of such systems that are capable of generating verifiably persuasiv…

2025

SPRO: Improving Image Generation via Self-Play

NeurIPS 2025poster

Recent advances in diffusion models have dramatically improved image fidelity and diversity. However, aligning these models with nuanced human preferences -such as aesthetics, engagement, and subjective appeal remains a key challenge due to the scarcity of large-scale human annotations. Collecting s…

Cited by 0SourceScholar
2025

Teaching Human Behavior Improves Content Understanding Abilities Of VLMs

ICLR 2025poster

Communication is defined as "*Who* says *what* to *whom* with *what* effect." A message from a communicator generates downstream receiver effects, also known as behavior. Receiver behavior, being a downstream effect of the message, carries rich signals about it. Even after carrying signals about the…

2022

MINIMAL: Mining Models for Universal Adversarial Triggers

AAAI 2022technical

It is well known that natural language models are vulnerable to adversarial attacks, which are mostly input-specific in nature. Recently, it has been shown that there also exist input-agnostic attacks in NLP models, called universal adversarial triggers. However, existing methods to craft universal…

Cited by 5SourcePDFScholar
2021

LIFI: Towards Linguistically Informed Frame Interpolation

ICASSP 2021accepted

Here we explore the problem of speech video interpolation. With close to 70% of web traffic, such content today forms the primary form of online communication and entertainment. Despite high performance on conventional metrics like MSE, PSNR, and SSIM, we find that the state-of-the-art frame interpo…

Cited by 0SourceScholar