← Search

Helen Yannakoudakis

10 accepted papers

2025

A Survey of Cognitive Distortion Detection and Classification in NLP

EMNLP 2025

As interest grows in applying natural language processing (NLP) techniques to mental health, an expanding body of work explores the automatic detection and classification of cognitive distortions (CDs). CDs are habitual patterns of negatively biased or flawed thinking that distort how people perceiv

2024

A (More) Realistic Evaluation Setup for Generalisation of Community Models on Malicious Content Detection

NAACL 2024findings

Community models for malicious content detection, which take into account the context from a social graph alongside the content itself, have shown remarkable performance on benchmark datasets. Yet, misinformation and hate speech continue to propagate on social media networks. This mismatch can be pa…

2024

Logging Keystrokes in Writing by English Learners

COLING 2024main

Essay writing is a skill commonly taught and practised in schools. The ability to write a fluent and persuasive essay is often a major component of formal assessment. In natural language processing and education technology we may work with essays in their final form, for example to carry out automat…

Cited by 5SourcePDFScholar
2024

Prompting open-source and commercial language models for grammatical error correction of English learner text

ACL 2024findings

Thanks to recent advances in generative AI, we are able to prompt large language models (LLMs) to produce texts which are fluent and grammatical. In addition, it has been shown that we can elicit attempts at grammatical error correction (GEC) from LLMs when prompted with ungrammatical input sentence…

Cited by 20SourcePDFScholar
2022

Meta-Learning for Fast Cross-Lingual Adaptation in Dependency Parsing

ACL 2022long

Meta-learning, or learning to learn, is a technique that can help to overcome resource scarcity in cross-lingual NLP problems, by enabling fast adaptation to new tasks. We apply model-agnostic meta-learning (MAML) to the task of cross-lingual dependency parsing. We train our model on a diverse set o…

2022

Scientific and Creative Analogies in Pretrained Language Models

EMNLP 2022finding

This paper examines the encoding of analogy in large-scale pretrained language models, such as BERT and GPT-2. Existing analogy datasets typically focus on a limited set of analogical relations, with a high similarity of the two domains between which the analogy holds. As a more realistic setup, we…

2021

Modeling Users and Online Communities for Abuse Detection: A Position on Ethics and Explainability

EMNLP 2021finding

Abuse on the Internet is an important societal problem of our time. Millions of Internet users face harassment, racism, personal attacks, and other types of abuse across various platforms. The psychological effects of abuse on individuals can be profound and lasting. Consequently, over the past few…

Cited by 11SourcePDFScholar
2021

Ruddit: Norms of Offensiveness for English Reddit Comments

ACL 2021long

On social media platforms, hateful and offensive language negatively impact the mental well-being of users and the participation of people from diverse backgrounds. Automatic methods to detect offensive language have largely relied on datasets with categorical labels. However, comments can vary in t…

2021

Towards a robust experimental framework and benchmark for lifelong language learning

NeurIPS 2021poster

In lifelong learning, a model learns different tasks sequentially throughout its lifetime. State-of-the-art deep learning models, however, struggle to generalize in this setting and suffer from catastrophic forgetting of old tasks when learning new ones. While a number of approaches have been develo…

Cited by 10SourceScholar