← Search

Mauro Cettolo

5 accepted papers

2024

Evaluating Automatic Subtitling: Correlating Post-editing Effort and Automatic Metrics

COLING 2024main

Systems that automatically generate subtitles from video are gradually entering subtitling workflows, both for supporting subtitlers and for accessibility purposes. Even though robust metrics are essential for evaluating the quality of automatically-generated subtitles and for estimating potential p…

2024

MOSEL: 950,000 Hours of Speech Data for Open-Source Speech Foundation Model Training on EU Languages

EMNLP 2024main

The rise of foundation models (FMs), coupled with regulatory efforts addressing their risks and impacts, has sparked significant interest in open-source models. However, existing speech FMs (SFMs) fall short of full compliance with the open-source principles, even if claimed otherwise, as no existin…

2024

SBAAM! Eliminating Transcript Dependency in Automatic Subtitling

ACL 2024long

Subtitling plays a crucial role in enhancing the accessibility of audiovisual content and encompasses three primary subtasks: translating spoken dialogue, segmenting translations into concise textual units, and estimating timestamps that govern their on-screen duration. Past attempts to automate thi…

2023

Integrating Language Models into Direct Speech Translation: An Inference-Time Solution to Control Gender Inflection

EMNLP 2023short main

When translating words referring to the speaker, speech translation (ST) systems should not resort to default masculine generics nor rely on potentially misleading vocal traits. Rather, they should assign gender according to the speakers' preference. The existing solutions to do so, though effectiv…

Cited by 0SourcecodeScholar
2021

Cascade versus Direct Speech Translation: Do the Differences Still Make a Difference?

ACL 2021long

Five years after the first published proofs of concept, direct approaches to speech translation (ST) are now competing with traditional cascade solutions. In light of this steady progress, can we claim that the performance gap between the two is closed? Starting from this question, we present a syst…