← Search

Wushour Silamu

5 accepted papers

2025

Improved Cross-Lingual Speaker Verification Using Speaker Sensitive Feature Guidance and Fine-grained Phonetic Information

ICASSP 2025accepted

Speaker verification performance significantly degrades when there exists a language mismatch between training and evaluation. Domain Adversarial Training (DAT) has shown to be effective in mitigating this gap by incorporating adversarial training with domain information (language id). Inspired by r…

Cited by 0SourceScholar
2025

Robust and Efficient Text-based Speech Editing using Noise Conditioning and Rectified Flow

ICASSP 2025accepted

Significant advancements have been made in text-based speech editing (TSE) for clear speech, but effectively editing the noise-contaminated speech remains a challenge. Background noise degrades the quality of generated speech, and edited speech that fails to maintain noise context consistency often…

Cited by 0SourceScholar
2024

Domain-Slot Aware Contrastive Learning for Improved Dialogue State Tracking

ICASSP 2024accepted

Large-scale pre-trained neural language model has facilitated to achieve the state-of-the-art performance on Dialogue State Tracking (DST) tasks. One of the existing works models the semantic correlation between the dialogue context and (domain, slot) pair encoded by BERT and make the prediction. De…

Cited by 0SourceScholar
2024

Fact-Aware Summarization with Contrastive Learning for Few-Shot Dialogue State Tracking

ICASSP 2024accepted

Dialogue state tracking (DST) is a crucial component of task-oriented dialogue systems, as it aims to accurately track the user’s goals throughout the dialogue history. However, DST models struggle with new domains due to limited annotated data, leading to poor performance. To solve this key challen…

Cited by 0SourceScholar