← Search

Jae Hyun Park

2 accepted papers

2024

SYNTHE-SEES: Face Based Text-to-Speech for Virtual Speaker

ICASSP 2024accepted

Recent virtual voice generation researches have limitations in that they results in low-quality voice and generate inconsistent voice from the same speaker’s different facial images. To handle this, we propose a facial encoder module for the pre-trained multi-speaker TTS system called SYNTHE-SEES, w…

Cited by 0SourceScholar
2021

HandFoldingNet: A 3D Hand Pose Estimation Network Using Multiscale-Feature Guided Folding of a 2D Hand Skeleton

ICCV 2021poster

With increasing applications of 3D hand pose estimation in various human-computer interaction applications, convolution neural networks (CNNs) based estimation models have been actively explored. However, the existing models require complex architectures or redundant computational resources to trade…

Cited by 56PDFcodeScholar