← Search

Hieu Minh Nguyen

2 accepted papers

2025

DarkBench: Benchmarking Dark Patterns in Large Language Models

ICLR 2025oral

We introduce DarkBench, a comprehensive benchmark for detecting dark design patterns—manipulative techniques that influence user behavior—in interactions with large language models (LLMs). Our benchmark comprises 660 prompts across six categories: brand bias, user retention, sycophancy, anthropomorp…

Cited by 34SourcePDFScholar
2024

“Global is Good, Local is Bad?”: Understanding Brand Bias in LLMs

EMNLP 2024main

Many recent studies have investigated social biases in LLMs but brand bias has received little attention. This research examines the biases exhibited by LLMs towards different brands, a significant concern given the widespread use of LLMs in affected use cases such as product recommendation and mark…

Cited by 3SourcePDFScholar