← Search

Benjamin Bergen

5 accepted papers

2025

Why do language models perform worse for morphologically complex languages?

COLING 2025main

Language models perform differently across languages. It has been previously suggested that morphological typology may explain some of this variability (Cotterell et al., 2018). We replicate previous analyses and find additional new evidence for a performance gap between agglutinative and fusional l…

2024

Correlations between Multilingual Language Model Geometry and Crosslingual Transfer Performance

COLING 2024main

A common approach to interpreting multilingual language models is to evaluate their internal representations. For example, studies have found that languages occupy distinct subspaces in the models’ representation spaces, and geometric distances between languages often reflect linguistic properties s…

Cited by 0SourcePDFScholar
2023

Rarely a problem? Language models exhibit inverse scaling in their predictions following few-type quantifiers

ACL 2023findings

How well do language models deal with quantification? In this study, we focus on ‘few’-type quantifiers, as in ‘few children like toys’, which might pose a particular challenge for language models because the sentence components with out the quantifier are likely to co-occur, and ‘few’-type quantifi…

Cited by 13SourcePDFScholar
2021

RAW-C: Relatedness of Ambiguous Words in Context (A New Lexical Resource for English)

ACL 2021long

Most words are ambiguous—-i.e., they convey distinct meanings in different contexts—-and even the meanings of unambiguous words are context-dependent. Both phenomena present a challenge for NLP. Recently, the advent of contextualized word embeddings has led to success on tasks involving lexical ambi…