← Search

Lucie Flek

19 accepted papers

2026

In-Training Defenses Against Emergent Misalignment in Language Models

ICML 2026poster

Fine‑tuning lets practitioners repurpose aligned large language models (LLMs) for new domains, yet recent work reveals emergent misalignment (EMA): Even a small, domain‑specific fine‑tune can induce harmful behaviors far outside the target domain. Even in the case where model weights are hidden behi…

Cited by 0SourceScholar
2026

PerSpectra: A Scalable and Configurable Pluralist Benchmark of Perspectives from Arguments

ICLR 2026poster

Pluralism, the capacity to engage with diverse perspectives without collapsing them into a single viewpoint, is critical for developing large language models that faithfully reflect human heterogeneity. Yet this characteristic has not been carefully examined within the LLM research community and rem…

Cited by 1SourcecodeScholar
2025

Explainable Hallucination through Natural Language Inference Mapping

ACL 2025finding

Large language models (LLMs) often generate hallucinated content, making it crucial to identify and quantify inconsistencies in their outputs. We introduce HaluMap, a post-hoc framework that detects hallucinations by mapping entailment and contradiction relations between source inputs and generated…

2025

Multi-Hop Reasoning for Question Answering with Hyperbolic Representations

ACL 2025finding

Hyperbolic representations are effective in modeling knowledge graph data which is prevalently used to facilitate multi-hop reasoning. However, a rigorous and detailed comparison of the two spaces for this task is lacking. In this paper, through a simple integration of hyperbolic representations wit…

2025

The Practical Impacts of Theoretical Constructs on Empathy Modeling

EMNLP 2025

Conceptual operationalizations of empathy in NLP are varied, with some having specific behaviors and properties, while others are more abstract. How these variations relate to one another and capture properties of empathy observable in text remains unclear. To provide insight into this, we analyze t

Cited by 0SourcePDFScholar
2025

USDC: A Dataset of  ̲User  ̲Stance and  ̲Dogmatism in Long  ̲Conversations

ACL 2025finding

Analyzing user opinion changes in long conversation threads is extremely critical for applications like enhanced personalization, market research, political campaigns, customer service, targeted advertising, and content moderation. Unfortunately, previous studies on stance and dogmatism in user conv…

2024

Appraisal Framework for Clinical Empathy: A Novel Application to Breaking Bad News Conversations

COLING 2024main

Empathy is essential in healthcare communication. We introduce an annotation approach that draws on well-established frameworks for clinical empathy and breaking bad news (BBN) conversations for considering the interactive dynamics of discourse relations. We construct Empathy in BBNs, a span-relatio…

2024

Corpus Considerations for Annotator Modeling and Scaling

NAACL 2024long

Recent trends in natural language processing research and annotation tasks affirm a paradigm shift from the traditional reliance on a single ground truth to a focus on individual perspectives, particularly in subjective tasks. In scenarios where annotation tasks are meant to encompass diversity, mod…

2024

DeFaktS: A German Dataset for Fine-Grained Disinformation Detection through Social Media Framing

COLING 2024main

In today’s rapidly evolving digital age, disinformation poses a significant threat to public sentiment and socio-political dynamics. To address this, we introduce a new dataset “DeFaktS”, designed to understand and counter disinformation within German media. Distinctively curated across various news…

2024

LeadEmpathy: An Expert Annotated German Dataset of Empathy in Written Leadership Communication

COLING 2024main

Empathetic leadership communication plays a pivotal role in modern workplaces as it is associated with a wide range of positive individual and organizational outcomes. This paper introduces LeadEmpathy, an innovative expert-annotated German dataset for modeling empathy in written leadership communic…

2024

The Impact of Differential Privacy on Group Disparity Mitigation

NAACL 2024findings

The performance cost of differential privacy has, for some applications, been shown to be higher for minority groups; fairness, conversely, has been shown to disproportionally compromise the privacy of members of such groups. Most work in this area has been restricted to computer vision and risk ass…

2022

A Critical Reflection and Forward Perspective on Empathy and Natural Language Processing

EMNLP 2022finding

We review the state of research on empathy in natural language processing and identify the following issues: (1) empathy definitions are absent or abstract, which (2) leads to low construct validity and reproducibility. Moreover, (3) emotional empathy is overemphasized, skewing our focus to a narrow…

Cited by 26SourcePDFScholar
2022

DMix: Adaptive Distance-aware Interpolative Mixup

ACL 2022short

Interpolation-based regularisation methods such as Mixup, which generate virtual training samples, have proven to be effective for various tasks and modalities. We extend Mixup and propose DMix, an adaptive distance-aware interpolative Mixup that selects samples based on their diversity in the embed…

2022

Mitigating Toxic Degeneration with Empathetic Data: Exploring the Relationship Between Toxicity and Empathy

NAACL 2022long

Large pre-trained neural language models have supported the effectiveness of many NLP tasks, yet are still prone to generating toxic language hindering the safety of their use. Using empathetic data, we improve over recent work on controllable text generation that aims to reduce the toxicity of gene…

Cited by 17SourcePDFScholar
2022

Unifying Data Perspectivism and Personalization: An Application to Social Norms

EMNLP 2022main

Instead of using a single ground truth for language processing tasks, several recent studies have examined how to represent and predict the labels of the set of annotators. However, often little or no information about annotators is known, or the set of annotators is small. In this work, we examine…

2021

HypMix: Hyperbolic Interpolative Data Augmentation

EMNLP 2021main

Interpolation-based regularisation methods for data augmentation have proven to be effective for various tasks and modalities. These methods involve performing mathematical operations over the raw input samples or their latent states representations - vectors that often possess complex hierarchical…

2021

Suicide Ideation Detection via Social and Temporal User Representations using Hyperbolic Learning

NAACL 2021long

Recent psychological studies indicate that individuals exhibiting suicidal ideation increasingly turn to social media rather than mental health practitioners. Personally contextualizing the buildup of such ideation is critical for accurate identification of users at risk. In this work, we propose a…

Cited by 56SourcePDFScholar