← Search

Antske Fokkens

11 accepted papers

2025

DefVerify: Do Hate Speech Models Reflect Their Dataset’s Definition?

COLING 2025main

When building a predictive model, it is often difficult to ensure that application-specific requirements are encoded by the model that will eventually be deployed. Consider researchers working on hate speech detection. They will have an idea of what is considered hate speech, but building a model th…

2025

Engagement-driven Persona Prompting for Rewriting News Tweets

COLING 2025main

Text style transfer is a challenging research task which modifies the linguistic style of a given text to meet pre-set objectives such as making the text simpler or more accessible. Though large language models have been found to give promising results, text rewriting to improve audience engagement…

2025

Improving Causal Interventions in Amnesic Probing with Mean Projection or LEACE

ACL 2025finding

Amnesic probing is a technique used to examine the influence of specific linguistic information on the behaviour of a model. This involves identifying and removing the relevant information and then assessing whether the model’s performance on the main task changes. If the removed information is rele…

Cited by 0SourcePDFScholar
2024

Investigating the Robustness of Modelling Decisions for Few-Shot Cross-Topic Stance Detection: A Preregistered Study

COLING 2024main

For a viewpoint-diverse news recommender, identifying whether two news articles express the same viewpoint is essential. One way to determine “same or different” viewpoint is stance detection. In this paper, we investigate the robustness of operationalization choices for few-shot stance detection, w…

2024

The Role of Syntactic Span Preferences in Post-Hoc Explanation Disagreement

COLING 2024main

Post-hoc explanation methods are an important tool for increasing model transparency for users. Unfortunately, the currently used methods for attributing token importance often yield diverging patterns. In this work, we study potential sources of disagreement across methods from a linguistic perspec…

2023

Dynamic Top-k Estimation Consolidates Disagreement between Feature Attribution Methods

EMNLP 2023short main

Feature attribution scores are used for explaining the prediction of a text classifier to users by highlighting a k number of tokens. In this work, we propose a way to determine the number of optimal k tokens that should be displayed from sequential properties of the attribution scores. Our approach…

Cited by 0SourceScholar
2023

Methodological Insights in Detecting Subtle Semantic Shifts with Contextualized and Static Language Models

EMNLP 2023long findings

In this paper, we investigate automatic detection of subtle semantic shifts between social communities of different political convictions in Dutch and English. We perform a methodological study comparing methods using static and contextualized language models. We investigate the impact of specializi…

Cited by 0SourceScholar
2022

Better Hit the Nail on the Head than Beat around the Bush: Removing Protected Attributes with a Single Projection

EMNLP 2022main

Bias elimination and recent probing studies attempt to remove specific information from embedding spaces. Here it is important to remove as much of the target information as possible, while preserving any other information present. INLP is a popular recent method which removes specific information t…

Cited by 10SourcePDFScholar
2022

Dealing with Abbreviations in the Slovenian Biographical Lexicon

EMNLP 2022main

Abbreviations present a significant challenge for NLP systems because they cause tokenization and out-of-vocabulary errors. They can also make the text less readable, especially in reference printed books, where they are extensively used. Abbreviations are especially problematic in low-resource sett…

2021

Challenging distributional models with a conceptual network of philosophical terms

NAACL 2021long

Computational linguistic research on language change through distributional semantic (DS) models has inspired researchers from fields such as philosophy and literary studies, who use these methods for the exploration and comparison of comparatively small datasets traditionally analyzed by close read…

2020

Would you describe a leopard as yellow? Evaluating crowd-annotations with justified and informative disagreement

COLING 2020main

Semantic annotation tasks contain ambiguity and vagueness and require varying degrees of world knowledge. Disagreement is an important indication of these phenomena. Most traditional evaluation methods, however, critically hinge upon the notion of inter-annotator agreement. While alternative framewo…