← Search

Leo Wanner

6 accepted papers

2025

Exploring morphology-aware tokenization: A case study on Spanish language modeling

EMNLP 2025

This paper investigates to what extent the integration of morphological information can improve subword tokenization and thus also language modeling performance. We focus on Spanish, a language with fusional morphology, where subword segmentation can benefit from linguistic structure. Instead of rel

2024

GPT-HateCheck: Can LLMs Write Better Functional Tests for Hate Speech Detection?

COLING 2024main

Online hate detection suffers from biases incurred in data sampling, annotation, and model pre-training. Therefore, measuring the averaged performance over all examples in held-out test data is inadequate. Instead, we must identify specific model weaknesses and be informed when it is more likely to…

2024

Using Large Language Models and Recruiter Expertise for Optimized Multilingual Job Offer – Applicant CV Matching

IJCAI 2024poster

In the context of the increasingly globalised economy and labour market, recruitment agencies face the challenge to deal with a magnitude of job offers and job applications written in a variety of languages, formats, and styles. Quite often, this leads to a suboptimal evaluation of the CVs of job se…

2022

Directions for NLP Practices Applied to Online Hate Speech Detection

EMNLP 2022main

Addressing hate speech in online spaces has been conceptualized as a classification task that uses Natural Language Processing (NLP) techniques. Through this conceptualization, the hate speech detection task has relied on common conventions and practices from NLP. For instance, inter-annotator agree…

Cited by 30SourcePDFScholar
2021

How much pretraining data do language models need to learn syntax?

EMNLP 2021main

Transformers-based pretrained language models achieve outstanding results in many well-known NLU benchmarks. However, while pretraining methods are very convenient, they are expensive in terms of time and resources. This calls for a study of the impact of pretraining data size on the knowledge of th…

Cited by 42SourcePDFScholar