← Search

Tran Huy Dat

4 accepted papers

2025

Automatic Speech Recognition and Spoken Language Understanding of Maritime Radio Communications: A case study with Singapore data

ICASSP 2025accepted

Speech communication has been a major part of maritime transportation, particularly in Vessel Traffic Management (VTM). Automatic Speech Recognition (ASR) and Spoken Language Understanding (SLU) of VTM communications provide shipping traffic information in digital forms which improves the productivi…

Cited by 0SourceScholar
2019

Embedding Physical Augmentation and Wavelet Scattering Transform to Generative Adversarial Networks for Audio Classification with Limited Training Resources

ICASSP 2019accepted

This paper addresses audio classification with limited training resources. We first investigate different types of data augmentation including physical modeling, wavelet scattering transform and Generative Adversarial Networks (GAN). We than propose a novel GAN method to embed physical augmentation…

Cited by 0SourceScholar
2016

A comparative study of multi-channel processing methods for noisy automatic speech recognition in urban environments

ICASSP 2016accepted

For the distant speech recognition, the multi-channel processing has been proven to significantly improve the ASR performances compared to the single channel approaches. However, there is very little work has done to provide a comparative evaluation of the approaches, particularly with the modern De…

Cited by 0SourceScholar
2015

Combining robust spike coding with spiking neural networks for sound event classification

ICASSP 2015accepted

This paper proposes a novel biologically inspired method for sound event classification which combines spike coding with a spiking neural network (SNN). Our spike coding extracts keypoints that represent the local maxima components of the sound spectrogram, and are encoded based on their local time-…

Cited by 0SourceScholar