← Search

Mohammad Javad Dousti

5 accepted papers

2025

Dynamic Jointly Batch Selection for Data Efficient Machine Translation Fine-Tuning

EMNLP 2025

Data quality and its effective selection are fundamental to improving the performance of machine translation models, serving as cornerstones for achieving robust and reliable translation systems. This paper presents a data selection methodology specifically designed for fine-tuning machine translati

Cited by 0SourcePDFScholar
2024

EPOQUE: An English-Persian Quality Estimation Dataset

COLING 2024main

Translation quality estimation (QE) is an important component in real-world machine translation applications. Unfortunately, human labeled QE datasets, which play an important role in developing and assessing QE models, are only available for limited language pairs. In this paper, we present the fir…

2024

Esposito: An English-Persian Scientific Parallel Corpus for Machine Translation

COLING 2024main

Neural machine translation requires large number of parallel sentences along with in-domain parallel data to attain best results. Nevertheless, no scientific parallel corpus for English-Persian language pair is available. In this paper, a parallel corpus called Esposito is introduced, which contains…

Cited by 1SourcePDFScholar
2023

PMI-Align: Word Alignment With Point-Wise Mutual Information Without Requiring Parallel Training Data

ACL 2023findings

Word alignment has many applications including cross-lingual annotation projection, bilingual lexicon extraction, and the evaluation or analysis of translation outputs. Recent studies show that using contextualized embeddings from pre-trained multilingual language models could give us high quality w…

2021

Streaming Simultaneous Speech Translation with Augmented Memory Transformer

ICASSP 2021accepted

Transformer-based models have achieved state-of-the-art performance on speech translation tasks. However, the model architecture is not efficient enough for streaming scenarios since self-attention is computed over an entire input sequence and the computational cost grows quadratically with the leng…

Cited by 0SourceScholar