← Search

Rasmi Elasmar

1 accepted papers

2026

Multi-turn Evaluation of Anthropomorphic Behaviours in Large Language Models

ICLR 2026poster

The tendency of users to anthropomorphise large language models (LLMs) is of growing societal interest. Here, we present AnthroBench: a novel empirical method and tool for evaluating anthropomorphic LLM behaviours in realistic settings. Our work introduces three key advances; first, we develop a mul…

Cited by 0SourcecodeScholar