← Search

Tal August

8 accepted papers

2025

Tree-of-Debate: Multi-Persona Debate Trees Elicit Critical Thinking for Scientific Comparative Analysis

ACL 2025long

With the exponential growth of research facilitated by modern technology and improved accessibility, scientific discoveries have become increasingly fragmented within and across fields. This makes it challenging to assess the significance, novelty, incremental findings, and equivalent ideas between…

2024

APPLS: Evaluating Evaluation Metrics for Plain Language Summarization

EMNLP 2024main

While there has been significant development of models for Plain Language Summarization (PLS), evaluation remains a challenge. PLS lacks a dedicated assessment metric, and the suitability of text generation evaluation metrics is unclear due to the unique transformations involved (e.g., adding backgr…

2024

Leveraging Large Language Models for Learning Complex Legal Concepts through Storytelling

ACL 2024long

Making legal knowledge accessible to non-experts is crucial for enhancing general legal literacy and encouraging civic participation in democracy. However, legal documents are often challenging to understand for people without legal backgrounds. In this paper, we present a novel application of large…

2024

MathFish: Evaluating Language Model Math Reasoning via Grounding in Educational Curricula

EMNLP 2024finding

To ensure that math curriculum is grade-appropriate and aligns with critical skills or concepts in accordance with educational standards, pedagogical experts can spend months carefully reviewing published math problems. Drawing inspiration from this process, our work presents a novel angle for evalu…

2024

Personalized Jargon Identification for Enhanced Interdisciplinary Communication

NAACL 2024long

Scientific jargon can confuse researchers when they read materials from other domains. Identifying and translating jargon for individual researchers could speed up research, but current methods of jargon identification mainly use corpus-level familiarity indicators rather than modeling researcher-sp…

2022

Generating Scientific Definitions with Controllable Complexity

ACL 2022long

Unfamiliar terminology and complex language can present barriers to understanding science. Natural language processing stands to help address these issues by automatically defining unfamiliar terms. We introduce a new task and dataset for defining scientific terms and controlling the complexity of g…

2021

All That’s ‘Human’ Is Not Gold: Evaluating Human Evaluation of Generated Text

ACL 2021long

Human evaluations are typically considered the gold standard in natural language generation, but as models’ fluency improves, how well can evaluators detect and judge machine-generated text? We run a study assessing non-experts’ ability to distinguish between human- and machine-authored text (GPT2 a…