← Search

Dong-Yan Huang

3 accepted papers

2025

MHSDB: A Comprehensive Benchmark for Multimodal Humor and Sarcasm Detection Leveraging Foundation Models

ICASSP 2025accepted

Understanding multimodal humor and sarcasm detection remains a key challenge in artificial intelligence. Despite recent advances, inconsistencies in feature extraction, evaluation methods, and experimental setups have hindered fair comparisons across different approaches. To address this issue, we p…

Cited by 0SourceScholar
2016

Combining multiple kernel models for automatic intelligibility detection of pathological speech

ICASSP 2016accepted

Automatic detection of pathological voice is a challenging task in speech processing. Appropriate acoustic cues of voice can be used to differentiate between normal voices and pathological voices. We propose a method to represent each speech utterance using three types of speech signal representatio…

Cited by 4SourceScholar
2016

Exemplar-based sparse representation of timbre and prosody for voice conversion

ICASSP 2016accepted

Voice conversion (VC) aims to make one speaker (source) to sound like spoken by another speaker (target) without changing the language content. Most of the state-of-the-art voice conversion systems focus only on timbre conversion. However, the speaker identity is characterized by the source-related…

Cited by 0SourceScholar