← Search

Stéphane Aroca-Ouellette

3 accepted papers

2025

Aligning LLMs by Predicting Preferences from User Writing Samples

ICML 2025poster

Accommodating human preferences is essential for creating aligned LLM agents that deliver personalized and effective interactions. Recent work has shown the potential for LLMs acting as writing agents to infer a description of user preferences. Agent alignment then comes from conditioning on the inf…

2025

Implicitly Aligning Humans and Autonomous Agents through Shared Task Abstractions

IJCAI 2025

In collaborative tasks, autonomous agents fall short of humans in their capability to quickly adapt to new and unfamiliar teammates. We posit that a limiting factor for zero-shot coordination is the lack of shared task abstractions, a mechanism humans rely on to implicitly align with teammates. To a

2021

The World of an Octopus: How Reporting Bias Influences a Language Model’s Perception of Color

EMNLP 2021main

Recent work has raised concerns about the inherent limitations of text-only pretraining. In this paper, we first demonstrate that reporting bias, the tendency of people to not state the obvious, is one of the causes of this limitation, and then investigate to what extent multimodal training can miti…