← Search

Ariel Goldstein

6 accepted papers

2025

Can LLMs Learn Macroeconomic Narratives from Social Media?

NAACL 2025findings

This study empirically tests the Narrative Economics hypothesis, which posits that narratives (ideas that are spread virally and affect public beliefs) can influence economic fluctuations. We introduce two curated datasets containing posts from X (formerly Twitter) which capture economy-related narr…

Cited by 6SourcePDFScholar
2025

Confidence Improves Self-Consistency in LLMs

ACL 2025finding

Self-consistency decoding enhances LLMs’ performance on reasoning tasks by sampling diverse reasoning paths and selecting the most frequent answer. However, it is computationally expensive, as sampling many of these (lengthy) paths is required to increase the chances that the correct answer emerges…

Cited by 0SourcePDFScholar
2025

Looking Beyond the Top-1: Transformers Determine Top Tokens in Order

ICML 2025poster

Uncovering the inner mechanisms of Transformer models offers insights into how they process and represent information. In this work, we analyze the computation performed by Transformers in the layers after the top-1 prediction remains fixed, known as the “saturation event”. We expand this concept to…

2024

Do Zombies Understand? A Choose-Your-Own-Adventure Exploration of Machine Cognition

ACL 2024findings

Recent advances in LLMs have sparked a debate on whether they understand text. In this position paper, we argue that opponents in this debate hold different definitions for understanding, and particularly differ in their view on the role of consciousness. To substantiate this claim, we propose a tho…

Cited by 1SourcePDFScholar
2023

Decoding Stumpers: Large Language Models vs. Human Problem-Solvers

EMNLP 2023short findings

This paper investigates the problem-solving capabilities of Large Language Models (LLMs) by evaluating their performance on stumpers, unique single-step intuition problems that pose challenges for human solvers but are easily verifiable. We compare the performance of four state-of-the-art LLMs (Davi…

Cited by 0SourceScholar