← Search

Javier Iranzo-Sánchez

6 accepted papers

2025

Going Beyond Your Expectations in Latency Metrics for Simultaneous Speech Translation

ACL 2025finding

Current evaluation practices in Simultaneous Speech Translation (SimulST) systems typically involve segmenting the input audio and corresponding translations, calculating quality and latency metrics for each segment, and averaging the results. Although this approach may provide a reliable estimation…

Cited by 0SourcePDFScholar
2025

Maintaining Prosodic Consistency in Automatic Dubbing for Better Isochrony

ICASSP 2025accepted

One of the major requirements of automatic dubbing, as an extension of speech-to-speech translation, is isochrony. It refers to the fact that dubbed speech should line up in time with the original speech, matching the phrase-pause arrangement of the source utterance. To achieve this, we explore and…

Cited by 0SourceScholar
2022

From Simultaneous to Streaming Machine Translation by Leveraging Streaming History

ACL 2022long

Simultaneous Machine Translation is the task of incrementally translating an input sentence before it is fully available. Currently, simultaneous translation is carried out by translating each sentence independently of the previously translated text. More generally, Streaming MT can be understood as…

Cited by 11SourcePDFScholar
2021

Stream-level Latency Evaluation for Simultaneous Machine Translation

EMNLP 2021finding

Simultaneous machine translation has recently gained traction thanks to significant quality improvements and the advent of streaming applications. Simultaneous translation systems need to find a trade-off between translation quality and response time, and with this purpose multiple latency measures…

2020

Europarl-ST: A Multilingual Corpus for Speech Translation of Parliamentary Debates

ICASSP 2020accepted

Current research into spoken language translation (SLT), or speech-to-text translation, is often hampered by the lack of specific data resources for this task, as currently available SLT datasets are restricted to a limited set of language pairs. In this paper we present Europarl-ST, a novel multili…

Cited by 0SourceScholar
2020

LSTM-Based One-Pass Decoder for Low-Latency Streaming

ICASSP 2020accepted

Current state-of-the-art models based on Long-Short Term Memory (LSTM) networks have been extensively used in ASR to improve performance. However, using LSTMs under a streaming setup is not straightforward due to real-time constraints. In this paper we present a novel streaming decoder that includes…

Cited by 0SourceScholar