← Search

Samuel Joseph Amouyal

3 accepted papers

2025

When the LM misunderstood the human chuckled: Analyzing garden path effects in humans and language models

ACL 2025long

Modern Large Language Models (LLMs) have shown human-like abilities in many language tasks, sparking interest in comparing LLMs’ and humans’ language processing. In this paper, we try to answer two questions: 1. What makes garden-path sentences hard to understand for humans? 2. Do the same reasons m…

Cited by 0SourcePDFScholar
2024

AssistantBench: Can Web Agents Solve Realistic and Time-Consuming Tasks?

EMNLP 2024main

Language agents, built on top of language models (LMs), are systems that can interact with complex environments, such as the open web. In this work, we examine whether such agents can perform realistic and time-consuming tasks on the web, e.g., monitoring real-estate markets or locating relevant nea…

Cited by 13SourcePDFScholar
2024

STEER: Assessing the Economic Rationality of Large Language Models

ICML 2024poster

There is increasing interest in using LLMs as decision-making "agents". Doing so includes many degrees of freedom: which model should be used; how should it be prompted; should it be asked to introspect, conduct chain-of-thought reasoning, etc? Settling these questions---and more broadly, determinin…

Cited by 16SourcePDFScholar