← Search

Seunghee Han

5 accepted papers

2025

A Domain-Specific Multilingual Speech Translation Corpus via Simultaneous Interpretation

ICASSP 2025accepted

This paper presents a novel multilingual speech translation corpus for complex, domain-specific content in Korean, English, Spanish, and Japanese. The corpus contains 4,000 hours of parallel speech, including 1,000 hours of Korean audio with simultaneous sight interpretations in the other three lang…

Cited by 0SourceScholar
2025

Personalized Lip Reading: Adapting to Your Unique Lip Movements with Vision and Language

AAAI 2025technical

Lip reading aims to predict spoken language by analyzing lip movements. Despite advancements in lip reading technologies, performance degrades when models are applied to unseen speakers due to their sensitivity to variations in visual information such as lip appearances. To address this challenge, s…

2024

Constructing Korean Learners’ L2 Speech Corpus of Seven Languages for Automatic Pronunciation Assessment

COLING 2024main

Multilingual L2 speech corpora for developing automatic speech assessment are currently available, but they lack comprehensive annotations of L2 speech from non-native speakers of various languages. This study introduces the methodology of designing a Korean learners’ L2 speech corpus of seven langu…

Cited by 2SourcePDFScholar
2024

Persona Extraction Through Semantic Similarity for Emotional Support Conversation Generation

ICASSP 2024accepted

Providing emotional support through dialogue systems is becoming increasingly important in today’s world, as it can support both mental health and social interactions in many conversation scenarios. Previous works have shown that using persona is effective for generating empathetic and supportive re…

Cited by 0SourceScholar
2024

Where Visual Speech Meets Language: VSP-LLM Framework for Efficient and Context-Aware Visual Speech Processing

EMNLP 2024finding

In visual speech processing, context modeling capability is one of the most important requirements due to the ambiguous nature of lip movements. For example, homophenes, words that share identical lip movements but produce different sounds, can be distinguished by considering the context. In this pa…