← Search

Andy W. H. Khong

20 accepted papers

2025

Contactless Vital Sign Monitoring for Multiple People Using a Millimeter-wave MIMO Radar

ICASSP 2025accepted

Radar technology offers much appeal for contactless vital sign monitoring. While most radar-based approaches achieve reasonable performance for single-person scenarios, they suffer from inaccurate vital sign estimates for multiple individuals, especially when the subjects occupy the same range bin.…

Cited by 0SourceScholar
2025

SMARTMiner: Extracting and Evaluating SMART Goals from Low-Resource Health Coaching Notes

EMNLP 2025

We present SMARTMiner, a framework for extracting and evaluating specific, measurable, attainable, relevant, time-bound (SMART) goals from unstructured health coaching (HC) notes. Developed in response to challenges observed during a clinical trial, the SMARTMiner achieves two tasks: (i) extracting

2025

ViKIENet: Towards Efficient 3D Object Detection with Virtual Key Instance Enhanced Network

CVPR 2025poster

The sparsity of point clouds and inadequacy of semantic information pose challenges to current LiDAR-only 3D object detection methods. Recent methods alleviate these challenges by converting RGB images into virtual points via depth completion to be fused with LiDAR points. Although these methods hav…

Cited by 0SourcePDFScholar
2024

Enhancing Code-Switching Speech Recognition With Interactive Language Biases

ICASSP 2024accepted

Languages usually switch within a multilingual speech signal, especially in a bilingual society. This phenomenon is referred to as code-switching (CS), making automatic speech recognition (ASR) challenging under a multilingual scenario. We propose to improve CS-ASR by biasing the hybrid CTC/attentio…

Cited by 30SourceScholar
2023

Improving Performance of Real-Time Full-Band Blind Packet-Loss Concealment with Predictive Network

ICASSP 2023accepted

Packet loss concealment (PLC) is a tool for enhancing speech degradation caused by poor network conditions or underflow/overflow in audio processing pipelines. We propose a real-time recurrent method that leverages previous outputs to mitigate artefact of lost packets without the prior knowledge of…

Cited by 0SourceScholar
2023

Reducing Language Confusion for Code-Switching Speech Recognition with Token-Level Language Diarization

ICASSP 2023accepted

Code-switching (CS) occurs when languages switch within a speech signal and leads to language confusion for automatic speech recognition (ASR). We address the problem of language confusion for improving CS-ASR from two perspectives: incorporating and disentangling language information. We incorporat…

Cited by 0SourceScholar
2022

Joint Source Localization and Association Through Overcomplete Representation Under Multipath Propagation Environment

ICASSP 2022accepted

This work addresses the source localization and association problem in a multipath propagation environment. By focusing on the limitation of the prior information in practical applications, we propose a target localization and association method based on iterative optimization with semi-unitary cons…

Cited by 0SourceScholar
2022

Multichannel Noise Reduction Using Dilated Multichannel U-Net and Pre-Trained Single-Channel Network

ICASSP 2022accepted

Pre-trained single-channel neural networks have become more prevalent for noise reduction in recent years. However, unlike their multichannel counterparts, these monoaural approaches do not exploit spatial information during the optimization process. Furthermore, while multichannel neural networks e…

Cited by 0SourceScholar
2022

Tunet: A Block-Online Bandwidth Extension Model Based On Transformers And Self-Supervised Pretraining

ICASSP 2022accepted

We introduce a block-online variant of the temporal feature-wise linear modulation (TFiLM) model to achieve bandwidth extension. The proposed architecture simplifies the UNet backbone of the TFiLM to reduce inference time and employs an efficient transformer at the bottleneck to alleviate performanc…

Cited by 0SourceScholar
2021

An Adaptive Non-Linear Process for Under-Determined Virtual Microphone Beamforming

ICASSP 2021accepted

Virtual microphone beamforming techniques are attractive for devices limited by space constraints. These techniques synthesize virtual microphone signals via interpolation algorithms. We propose to extend existing virtual microphone signal interpolation by employing an adaptive non-linear (ANL) proc…

Cited by 0SourceScholar
2021

Directional Sparse Filtering Using Weighted Lehmer Mean for Blind Separation of Unbalanced Speech Mixtures

ICASSP 2021accepted

In blind source separation of speech signals, the inherent imbalance in the source spectrum poses a challenge for methods that rely on single-source dominance for the estimation of the mixing matrix. We propose an algorithm based on the directional sparse filtering (DSF) framework that utilizes the…

Cited by 0SourceScholar
2018

Automatically Linking Digital Signal Processing Assessment Questions to Key Engineering Learning Outcomes

ICASSP 2018accepted

To deliver on the potential outcome-based teaching and learning holds for engineering education, it is important for engineering courses to provide students with different types of deliberate practice opportunities that align to the program's learning outcomes. Working from these requirements, we in…

Cited by 0SourceScholar
2018

Online Education Evaluation for Signal Processing Course Through Student Learning Pathways

ICASSP 2018accepted

Impact of online learning sequences to forecast course outcomes for an undergraduate digital signal processing (DSP) course is studied in this work. A multi-modal learning schema based on deep-learning techniques with learning sequences, psychometric measures, and personality traits as input feature…

Cited by 0SourceScholar
2017

Learning complex-valued latent filters with absolute cosine similarity

ICASSP 2017accepted

We propose a new sparse coding technique based on the power mean of phase-invariant cosine distances. Our approach is a generalization of sparse filtering and K-hyperlines clustering. It offers a better sparsity enforcer than the L <sub xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="htt…

Cited by 0SourceScholar
2017

On TOA estimation of vibration signals for localizing impacts on solid surfaces

ICASSP 2017accepted

We propose a TDOA-based algorithm for source localization on rigid surfaces. This allows the conversion of readily available large surfaces into touch interfaces using surface-mounted vibration sensors. To achieve this, we characterize the arrival of each sensor-received signal by the arrival times…

Cited by 0SourceScholar
2016

Source localization on solids utilizing logistic modeling of energy transition in vibration signals

ICASSP 2016accepted

We propose a new algorithm for source localization on rigid surfaces, which allows one to convert daily objects into human-computer touch interfaces using surface-mounted vibration sensors. This is achieved via estimating the time-difference-of-arrivals (TDOA) of the signals across the sensors. In t…

Cited by 0SourceScholar
2015

Multi-source direction-of-arrival estimation in a reverberant environment using single acoustic vector sensor

ICASSP 2015accepted

We address the problem of estimating direction-of-arrivals (DOAs) for multiple sound sources using a single acoustic vector sensor (AVS) in an enclosed room environment. It is well-known that multi-source DOA estimation in an enclosed environment is challenging due to room reverberation, environment…

Cited by 0SourceScholar
2015

Single-channel speech enhancement in a transient noise environment by exploiting speech harmonicity

ICASSP 2015accepted

This paper focuses on the problem of single-channel noise reduction in a transient noise environment for speech enhancement application. A typical speech enhancement algorithm requires an estimate of the noise statistics. However, the problem of noise estimation is challenging when the statistics of…

Cited by 0SourceScholar