← Search

B{\u{a}}nescu Ema-Ioana

1 accepted papers

2025

MoRoVoc: A Large Dataset for Geographical Variation Identification of the Spoken Romanian Language

EMNLP 2025

This paper introduces MoRoVoc, the largest dataset for analyzing the regional variation of spoken Romanian. It has more than 93 hours of audio and 88,192 audio samples, balanced between the Romanian language spoken in Romania and the Republic of Moldova. We further propose a multi-target adversarial

Cited by 0SourcePDFScholar