← Search

Ioan-Bogdan Iordache

7 accepted papers

2025

Friend or Foe? A Computational Investigation of Semantic False Friends across Romance Languages

EMNLP 2025

In this paper we present a comprehensive analysis of lexical semantic divergence between cognate words and borrowings in the Romance languages. We experiment with different algorithms for false friend detection including deceptive cognate and deceptive borrowings and correction and evaluate them sys

Cited by 0SourcePDFScholar
2024

It takes two to borrow: a donor and a recipient. Who’s who?

ACL 2024findings

We address the open problem of automatically identifying the direction of lexical borrowing, given word pairs in the donor and recipient languages. We propose strong benchmarks for this task, by applying a set of machine learning models. We extract and publicly release a comprehensive borrowings dat…

Cited by 1SourcePDFScholar
2024

Pater Incertus? There Is a Solution: Automatic Discrimination between Cognates and Borrowings for Romance Languages

COLING 2024main

Identifying the type of relationship between words (cognates, borrowings, inherited) provides a deeper insight into the history of a language and allows for a better characterization of language relatedness. In this paper, we propose a computational approach for discriminating between cognates and b…

2024

RoCode: A Dataset for Measuring Code Intelligence from Problem Definitions in Romanian

COLING 2024main

Recently, large language models (LLMs) have become increasingly powerful and have become capable of solving a plethora of tasks through proper instructions in natural language. However, the vast majority of testing suites assume that the instructions are written in English, the de facto prompting la…

2024

Verba volant, scripta volant? Don’t worry! There are computational solutions for protoword reconstruction

EMNLP 2024main

We introduce a new database of cognate words and etymons for the five main Romance languages, the most comprehensive one to date. We propose a strong benchmark for the automatic reconstruction of protowords for Romance languages, by applying a set of machine learning models and features on these dat…

2023

RoBoCoP: A Comprehensive ROmance BOrrowing COgnate Package and Benchmark for Multilingual Cognate Identification

EMNLP 2023long main

The identification of cognates is a fundamental process in historical linguistics, on which any further research is based. Even though there are several cognate databases for Romance languages, they are rather scattered, incomplete, noisy, contain unreliable information, or have uncertain availabili…

Cited by 10SourceScholar
2021

A Computational Exploration of Pejorative Language in Social Media

EMNLP 2021finding

In this paper we study pejorative language, an under-explored topic in computational linguistics. Unlike existing models of offensive language and hate speech, pejorative language manifests itself primarily at the lexical level, and describes a word that is used with a negative connotation, making i…

Cited by 22SourcePDFScholar