← Search

Ching Hua Lee

11 accepted papers

2026

HSG-12M: A Large-Scale Benchmark of Spatial Multigraphs from the Energy Spectra of Non-Hermitian Crystals

ICLR 2026poster

AI is transforming scientific research by revealing new ways to understand complex physical systems, but its impact remains constrained by the lack of large, high-quality domain-specific datasets. A rich, largely untapped resource lies in non-Hermitian quantum physics, where the energy spectra of cr…

Cited by 0SourcecodeScholar
2025

Better Exploiting Spatial Separability in Multichannel Speech Enhancement with an Align-and-Filter Network

ICASSP 2025accepted

Multichannel speech enhancement (SE) techniques combine multiple microphone signals to extract clean speech from noisy mixtures based on spatial filtering. As the target speech may come from arbitrary, unknown directions, current deep learning-based SE systems could suffer from performance bottlenec…

Cited by 0SourceScholar
2025

MIB: Mixed Information Bottleneck for Out-of-Distribution Keyword Spotting

ICASSP 2025accepted

Deep Keyword Spotting (KWS) systems continuously process audio streams to detect keywords. However, performance of deep neural networks degrade when the input data diverges from the training data; referred to as Out-of-Distribution (OOD) data problem. In this paper, we show performance degradation o…

Cited by 0SourceScholar
2024

An MVDR-Embedded U-Net Beamformer for Effective and Robust Multichannel Speech Enhancement

ICASSP 2024accepted

In multichannel speech enhancement (SE) systems, deep neural networks (DNNs) are often utilized to directly estimate the clean speech for effective beamforming. This approach, however, may not generalize adequately to new acoustic or noise conditions. Alternatively, DNNs can indirectly perform SE by…

Cited by 0SourceScholar
2024

End-To-End Personalized Cuff-Less Blood Pressure Monitoring Using ECG and PPG Signals

ICASSP 2024accepted

Cuffless blood pressure (BP) monitoring offers the potential for continuous, non-invasive healthcare but has been limited in adoption by existing models relying on handcrafted features from ECG and PPG signals. To overcome this, researchers have looked to deep learning. Along these lines, in this pa…

Cited by 0SourceScholar
2024

Leveraging Self-Supervised Speech Representations for Domain Adaptation in Speech Enhancement

ICASSP 2024accepted

Deep learning based speech enhancement (SE) approaches could suffer from performance degradation due to mismatch between training and testing environments. A realistic situation is that an SE model trained on parallel noisy-clean utterances from one environment, the source domain, may fail to perfor…

Cited by 0SourceScholar
2024

Zero-Shot Intent Classification Using a Semantic Similarity Aware Contrastive Loss and Large Language Model

ICASSP 2024accepted

Zero-shot systems can reduce the cost of collecting data and training in a new domain since they can work directly with the test data without further training. In this paper, we build zero-shot systems for intent classification, based on Semantic Similarity-aware Contrastive Loss (SSCL) that address…

Cited by 1SourceScholar
2023

A DNN Based Normalized Time-Frequency Weighted Criterion for Robust Wideband DoA Estimation

ICASSP 2023accepted

Deep neural networks (DNNs) have greatly benefited direction of arrival (DoA) estimation methods for speech source localization in noisy environments. However, their localization accuracy is still far from satisfactory due to the vulnerability to nonspeech interference. To improve the robustness aga…

Cited by 0SourceScholar
2023

Improved Mask-Based Neural Beamforming for Multichannel Speech Enhancement by Snapshot Matching Masking

ICASSP 2023accepted

In multichannel speech enhancement (SE), time-frequency (T-F) mask-based neural beamforming algorithms take advantage of deep neural networks to predict T-F masks that represent speech and noise dominance. The predicted masks are subsequently leveraged to estimate the speech and noise power spectral…

Cited by 0SourceScholar
2023

To Wake-Up or Not to Wake-Up: Reducing Keyword False Alarm by Successive Refinement

ICASSP 2023accepted

Keyword spotting systems continuously process audio streams to detect keywords. One of the most challenging tasks in designing such systems is to reduce False Alarm (FA) which happens when the system falsely registers a keyword despite the keyword not being uttered. In this paper, we propose a simpl…

Cited by 0SourceScholar
2020

SSGD: Sparsity-Promoting Stochastic Gradient Descent Algorithm for Unbiased Dnn Pruning

ICASSP 2020accepted

While deep neural networks (DNNs) have achieved state-of-the-art results in many fields, they are typically over-parameterized. Parameter redundancy, in turn, leads to inefficiency. Sparse signal recovery (SSR) techniques, on the other hand, find compact solutions to overcomplete linear problems. Th…

Cited by 6SourceScholar