← Search

Alexander Spangher

14 accepted papers

2026

WebDS: An End-to-End Benchmark for Web-based Data Science

ICLR 2026poster

Many real-world data science tasks involve complex web-based interactions: finding appropriate data available on the internet, synthesizing multimodal data from different locations, and producing summarized analyses. Existing web benchmarks often focus on simplistic interactions and often do not req…

Cited by 1SourcecodeScholar
2025

NewsInterview: a Dataset and a Playground to Evaluate LLMs’ Grounding Gap via Informational Interviews

ACL 2025long

Large Language Models (LLMs) have demonstrated impressive capabilities in generating coherent text but often struggle with grounding language and strategic dialogue. To address this gap, we focus on journalistic interviews, a domain rich in grounding communication and abundant in data. We curate a d…

Cited by 0SourcePDFScholar
2025

Spatial Layouts in News Homepages Capture Human Preferences

EMNLP 2025

Information prioritization plays an important role in the way we perceive and understand the world. Homepage layouts, which are daily and manually curated by expert human news editors, serve as a tangible proxy for this prioritization. In this work, we present NewsHomepages, a novel and massive data

2024

Are Large Language Models Capable of Generating Human-Level Narratives?

EMNLP 2024main

As daily reliance on large language models (LLMs) grows, assessing their generation quality is crucial to understanding how they might impact on our communications. This paper investigates the capability of LLMs in storytelling, focusing on narrative development and plot progression. We introduce a…

2024

Do LLMs Plan Like Human Writers? Comparing Journalist Coverage of Press Releases with LLMs

EMNLP 2024main

Journalists engage in multiple steps in the news writing process that depend on human creativity, like exploring different “angles” (i.e. the specific perspectives a reporter takes). These can potentially be aided by large language models (LLMs). By affecting planning decisions, such interventions c…

Cited by 9SourcePDFScholar
2024

Explaining Mixtures of Sources in News Articles

EMNLP 2024finding

Human writers plan, _then_ write. For large language models (LLMs) to play a role in longer-form article generation, we must understand the planning steps humans make before writing. We explore one kind of planning, source-selection in news, as a case-study for evaluating plans in long-form generati…

Cited by 2SourcePDFScholar
2024

LegalDiscourse: Interpreting When Laws Apply and To Whom

NAACL 2024long

While legal AI has made strides in recent years, it still struggles with basic legal concepts: _when_ does a law apply? _Who_ does it applies to? _What_ does it do? We take a _discourse_ approach to addressing these problems and introduce a novel taxonomy for span-and-relation parsing of legal texts…

Cited by 1SourcePDFScholar
2024

Stay on Topic with Classifier-Free Guidance

ICML 2024spotlight

Classifier-Free Guidance (CFG) has recently emerged in as a lightweight technique to encourage prompt-adherence in generations, yet has not yet been successfully applied to language modeling. In this work, we demonstrate across a wide array of benchmarks that CFG can be used broadly as an inference-…

Cited by 43SourcePDFScholar
2024

Tracking the Newsworthiness of Public Documents

ACL 2024long

Journalists regularly make decisions on whether or not to report stories, based on “news values”. In this work, we wish to explicitly model these decisions to explore _when_ and _why_ certain stories get press attention. This is challenging because very few labelled links between source documents an…

2023

Identifying Informational Sources in News Articles

EMNLP 2023long main

News articles are driven by the informational sources journalists use in reporting. Modeling when, how and why sources get used together in stories can help us better understand the information we consume and even help journalists with the task of producing it. In this work, we take steps toward thi…

Cited by 0SourcecodeScholar
2023

Learning Action Conditions from Instructional Manuals for Instruction Understanding

ACL 2023long

The ability to infer pre- and postconditions of an action is vital for comprehending complex instructions, and is essential for applications such as autonomous instruction-guided agents and assistive AI that supports humans to perform physical tasks. In this work, we propose a task dubbed action con…

2022

NewsEdits: A News Article Revision Dataset and a Novel Document-Level Reasoning Challenge

NAACL 2022long

News article revision histories provide clues to narrative and factual evolution in news articles. To facilitate analysis of this evolution, we present the first publicly available dataset of news revision histories, NewsEdits. Our dataset is large-scale and multilingual; it contains 1.2 million art…

2021

Multitask Semi-Supervised Learning for Class-Imbalanced Discourse Classification

EMNLP 2021main

As labeling schemas evolve over time, small differences can render datasets following older schemas unusable. This prevents researchers from building on top of previous annotation work and results in the existence, in discourse learning in particular, of many small class-imbalanced datasets. In this…

Cited by 28SourcePDFScholar