← Search

Wooseok Han

3 accepted papers

2025

Face-StyleSpeech: Enhancing Zero-shot Speech Synthesis from Face Images with Improved Face-to-Speech Mapping

ICASSP 2025accepted

Generating speech from a face image is crucial for developing virtual humans capable of interacting using their unique voices, without relying on pre-recorded human speech. In this paper, we propose Face-StyleSpeech, a zero-shot TextTo-Speech (TTS) synthesis model that generates natural speech condi…

Cited by 0SourceScholar
2025

Stable-TTS: Stable Speaker-Adaptive Text-to-Speech Synthesis via Prosody Prompting

ICASSP 2025accepted

Speaker-adaptive Text-to-Speech (TTS) synthesis has attracted considerable attention due to its broad range of applications, such as personalized voice assistant services. While several approaches have been proposed, they often exhibit high sensitivity to either the quantity or the quality of target…

Cited by 0SourceScholar
2023

Workspace Force/Acceleration Disturbance Observer for Precise and Safe Motion Control

IROS 2023poster

The use of impedance control has become widespread in applications requiring simultaneous position tracking and compliance in contact. However, disturbances such as friction and model uncertainties can adversely affect the performance of impedance-based motion control. The disturbance observer (DOB)…

Cited by 0SourceScholar