← Search

Hyojung Han

6 accepted papers

2025

Adapters for Altering LLM Vocabularies: What Languages Benefit the Most?

ICLR 2025poster

Vocabulary adaptation, which integrates new vocabulary into pre-trained language models, enables expansion to new languages and mitigates token over-fragmentation. However, existing approaches are limited by their reliance on heuristics or external embeddings. We propose VocADT, a novel method for v…

2024

XLAVS-R: Cross-Lingual Audio-Visual Speech Representation Learning for Noise-Robust Speech Perception

ACL 2024long

Speech recognition and translation systems perform poorly on noisy inputs, which are frequent in realistic environments. Augmenting these systems with visual signals has the potential to improve robustness to noise. However, audio-visual (AV) data is only available in limited amounts and for fewer l…

Cited by 6SourcePDFScholar
2023

Bridging Background Knowledge Gaps in Translation with Automatic Explicitation

EMNLP 2023long main

Translations help people understand content written in another language. However, even correct literal translations do not fulfill that goal when people lack the necessary background to understand them. Professional translators incorporate explicitations to explain the missing context by considering…

Cited by 0SourcecodeScholar
2022

SimQA: Detecting Simultaneous MT Errors through Word-by-Word Question Answering

EMNLP 2022main

Detractors of neural machine translation admit that while its translations are fluent, it sometimes gets key facts wrong. This is particularly important in simultaneous interpretation where translations have to be provided as fast as possible: before a sentence is complete. Yet, evaluations of simul…

Cited by 7SourcePDFScholar
2021

Task Aware Multi-Task Learning for Speech to Text Tasks

ICASSP 2021accepted

In general, the direct Speech-to-text translation (ST) is jointly trained with Automatic Speech Recognition (ASR), and Machine Translation (MT) tasks. However, the issues with the current joint learning strategies inhibit the knowledge transfer across these tasks. We propose a task modulation networ…

Cited by 0SourceScholar