← Search

Jakub Piskorski

8 accepted papers

2025

Entity Framing and Role Portrayal in the News

ACL 2025finding

We introduce a novel multilingual and hierarchical corpus annotated for entity framing and role portrayal in news articles. The dataset uses a unique taxonomy inspired by storytelling elements, comprising 22 fine-grained roles, or archetypes, nested within three main categories: protagonist, antagon…

Cited by 0SourcePDFScholar
2025

NarratEX Dataset: Explaining the Dominant Narratives in News Texts

EMNLP 2025

We present NarratEX, a dataset designed for the task of explaining the choice of the Dominant Narrative in a news article, and intended to support the research community in addressing challenges such as discourse polarization and propaganda detection. Our dataset comprises 1,056 news articles in fou

2025

PolyNarrative: A Multilingual, Multilabel, Multi-domain Dataset for Narrative Extraction from News Articles

ACL 2025long

We present polyNarrative, a new multilingual dataset of news articles, annotated for narratives. Narratives are overt or implicit claims, recurring across articles and languages, promoting a specific interpretation or viewpoint on an ongoing topic, often propagating mis/disinformation. We developed…

2024

Cross-lingual Named Entity Corpus for Slavic Languages

COLING 2024main

This paper presents a corpus manually annotated with named entities for six Slavic languages — Bulgarian, Czech, Polish, Slovenian, Russian, and Ukrainian. This work is the result of a series of shared tasks, conducted in 2017–2023 as a part of the Workshops on Slavic Natural Language Processing. Th…

2024

Exploring the Usability of Persuasion Techniques for Downstream Misinformation-related Classification Tasks

COLING 2024main

We systematically explore the predictive power of features derived from Persuasion Techniques detected in texts, for solving different tasks of interest for media analysis; notably: detecting mis/disinformation, fake news, propaganda, partisan news and conspiracy theories. Firstly, we propose a set…

Cited by 4SourcePDFScholar
2023

Holistic Inter-Annotator Agreement and Corpus Coherence Estimation in a Large-scale Multilingual Annotation Campaign

EMNLP 2023long main

In this paper we report on the complexity of persuasion technique annotation in the context of a large multilingual annotation campaign involving 6 languages and approximately 40 annotators. We highlight the techniques that appear to be difficult for humans to annotate and elaborate on our findings…

Cited by 0SourceScholar
2023

Multilingual Multifaceted Understanding of Online News in Terms of Genre, Framing, and Persuasion Techniques

ACL 2023long

We present a new multilingual multifacet dataset of news articles, each annotated for genre (objective news reporting vs. opinion vs. satire), framing (what key aspects are highlighted), and persuasion techniques (logical fallacies, emotional appeals, ad hominem attacks, etc.). The persuasion techni…

Cited by 33SourcePDFScholar
2020

New Benchmark Corpus and Models for Fine-grained Event Classification: To BERT or not to BERT?

COLING 2020main

We introduce a new set of benchmark datasets derived from ACLED data for fine-grained event classification and compare the performance of various state-of-the-art models on these datasets, including SVM based on TF-IDF character n-grams and neural context-free embeddings (GLOVE and FASTTEXT) as well…

Cited by 27SourcePDFScholar