← Search

Sung-Lin Yeh

5 accepted papers

2024

Revisiting Self-supervised Learning of Speech Representation from a Mutual Information Perspective

ICASSP 2024accepted

Existing studies on self-supervised speech representation learning have focused on developing new training methods and applying pre-trained models for different applications. However, the quality of these models is often measured by the performance of different downstream tasks. How well the represe…

Cited by 6SourceScholar
2023

Conditioning and Sampling in Variational Diffusion Models for Speech Super-Resolution

ICASSP 2023accepted

Recently, diffusion models (DMs) have been increasingly used in audio processing tasks, including speech super-resolution (SR), which aims to restore high-frequency content given low-resolution speech utterances. This is commonly achieved by conditioning the network of noise predictor with low-resol…

Cited by 0SourceScholar
2020

A Dialogical Emotion Decoder for Speech Motion Recognition in Spoken Dialog

ICASSP 2020accepted

Developing a robust emotion speech recognition (SER) system for human dialog is important in advancing conversational agent design. In this paper, we proposed a novel inference algorithm, a dialogical emotion decoding (DED) algorithm, that treats a dialog as a sequence and consecutively decode the e…

Cited by 26SourceScholar
2019

An Interaction-aware Attention Network for Speech Emotion Recognition in Spoken Dialogs

ICASSP 2019accepted

Obtaining robust speech emotion recognition (SER) in scenarios of spoken interactions is critical to the developments of next generation human-machine interface. Previous research has largely focused on performing SER by modeling each utterance of the dialog in isolation without considering the tran…

Cited by 0SourceScholar