← Search

Ori Ernst

12 accepted papers

2025

Improving the Calibration of Confidence Scores in Text Generation Using the Output Distribution’s Characteristics

ACL 2025short

Well-calibrated model confidence scores can improve the usefulness of text generation models. For example, users can be prompted to review predictions with low confidence scores, to prevent models from returning bad or potentially dangerous predictions. However, confidence metrics are not always wel…

Cited by 0SourcePDFScholar
2025

PreSumm: Predicting Summarization Performance Without Summarizing

ACL 2025finding

Despite recent advancements in automatic summarization, state-of-the-art models do not summarize all documents equally well, raising the question: why? While prior research has extensively analyzed summarization models, little attention has been given to the role of document characteristics in influ…

2025

Where Did That Come From? Sentence-Level Error-Tolerant Attribution

EMNLP 2025

Attribution is the process of identifying which parts of the source support a generated output. While attribution can help users verify content and assess faithfulness, existing task definitions typically exclude unsupported or hallucinated content leaving them unattributed, overlooking the potentia

2024

The Power of Summary-Source Alignments

ACL 2024findings

Multi-document summarization (MDS) is a challenging task, often decomposed to subtasks of salience and redundancy detection, followed by text generation.In this context, alignment of corresponding sentences between a reference summary and its source documents has been leveraged to generate training…

2023

OpenAsp: A Benchmark for Multi-document Open Aspect-based Summarization

EMNLP 2023long main

The performance of automatic summarization models has improved dramatically in recent years. Yet, there is still a gap in meeting specific information needs of users in real-world scenarios, particularly when a targeted summary is sought, such as in the useful aspect-based summarization setting targ…

Cited by 0SourcecodeScholar
2023

Re-Examining Summarization Evaluation across Multiple Quality Criteria

EMNLP 2023short findings

The common practice for assessing automatic evaluation metrics is to measure the correlation between their induced system rankings and those obtained by reliable human evaluation, where a higher correlation indicates a better metric. Yet, an intricate setting arises when an NLP task is evaluated by…

Cited by 0SourceScholar
2022

Extending Multi-Text Sentence Fusion Resources via Pyramid Annotations

NAACL 2022long

NLP models that process multiple texts often struggle in recognizing corresponding and salient information that is often differently phrased, and consolidating the redundancies across texts. To facilitate research of such challenges, the sentence fusion task was proposed, yet previous datasets for t…

2022

Proposition-Level Clustering for Multi-Document Summarization

NAACL 2022long

Text clustering methods were traditionally incorporated into multi-document summarization (MDS) as a means for coping with considerable information repetition. Particularly, clusters were leveraged to indicate information saliency as well as to avoid redundancy. Such prior methods focused on cluster…

2021

QA-Align: Representing Cross-Text Content Overlap by Aligning Question-Answer Propositions

EMNLP 2021main

Multi-text applications, such as multi-document summarization, are typically required to model redundancies across related texts. Current methods confronting consolidation struggle to fuse overlapping information. In order to explicitly represent content overlap, we propose to align predicate-argume…

2021

iFacetSum: Coreference-based Interactive Faceted Summarization for Multi-Document Exploration

EMNLP 2021system demonstrations

We introduce iFᴀᴄᴇᴛSᴜᴍ, a web application for exploring topical document collections. iFᴀᴄᴇᴛSᴜᴍ integrates interactive summarization together with faceted search, by providing a novel faceted navigation scheme that yields abstractive summaries for the user’s selections. This approach offers both a c…