← Search

Atharvan Dogra

1 accepted papers

2025

Language Models can Subtly Deceive Without Lying: A Case Study on Strategic Phrasing in Legislation

ACL 2025long

We explore the ability of large language models (LLMs) to engage in subtle deception through strategically phrasing and intentionally manipulating information. This harmful behavior can be hard to detect, unlike blatant lying or unintentional hallucination. We build a simple testbed mimicking a legi…