← Search

Hervé Bourlard

20 accepted papers

2021

Automatic And Perceptual Discrimination Between Dysarthria, Apraxia of Speech, and Neurotypical Speech

ICASSP 2021accepted

Automatic techniques in the context of motor speech disorders (MSDs) are typically two-class techniques aiming to discriminate between dysarthria and neurotypical speech or between dysarthria and apraxia of speech (AoS). Further, although such techniques are proposed to support the perceptual assess…

Cited by 0SourceScholar
2021

Automatic Dysarthric Speech Detection Exploiting Pairwise Distance-Based Convolutional Neural Networks

ICASSP 2021accepted

Automatic dysarthric speech detection can provide reliable and cost-effective computer-aided tools to assist the clinical diagnosis and management of dysarthria. In this paper we propose a novel automatic dysarthric speech detection approach based on analyses of pairwise distance matrices using conv…

Cited by 0SourceScholar
2021

Lattice-Free Mmi Adaptation of Self-Supervised Pretrained Acoustic Models

ICASSP 2021accepted

In this work, we propose lattice-free MMI (LFMMI) for supervised adaptation of self-supervised pretrained acoustic model. We pretrain a Transformer model on thousand hours of untranscribed Librispeech data followed by supervised adaptation with LFMMI on three different datasets. Our results show tha…

Cited by 0SourceScholar
2020

Incremental Semi-Supervised Learning for Multi-Genre Speech Recognition

ICASSP 2020accepted

In this work, we explore a data scheduling strategy for semi-supervised learning (SSL) for acoustic modeling in automatic speech recognition. The conventional approach uses a seed model trained with supervised data to automatically recognize the entire set of unlabeled (auxiliary) data to generate n…

Cited by 0SourceScholar
2020

Synthetic Speech References for Automatic Pathological Speech Intelligibility Assessment

ICASSP 2020accepted

Automatic pathological speech intelligibility measures are crucial to assist the clinical diagnosis and treatment of speech disorders. The recently proposed pathological short-time objective intelligibility (P-ESTOI) measure was shown to be very advantageous, yielding a high performance for several…

Cited by 0SourceScholar
2019

An End-to-end Network to Synthesize Intonation Using a Generalized Command Response Model

ICASSP 2019accepted

The generalized command response (GCR) model represents intonation as a superposition of muscle responses to spike command signals. We have previously shown that the spikes can be predicted by a two-stage system, consisting of a recurrent neural network and a post-processing procedure, but the respo…

Cited by 0SourceScholar
2019

Pathological Speech Intelligibility Assessment Based on the Short-time Objective Intelligibility Measure

ICASSP 2019accepted

Impaired speech intelligibility in motor speech disorders arising due to neurological diseases negatively affects the communication ability and quality of life of patients. Reliable and cost-effective measures to automatically assess speech intelligibility are necessary for the management of such di…

Cited by 42SourceScholar
2019

Super-gaussianity of Speech Spectral Coefficients as a Potential Biomarker for Dysarthric Speech Detection

ICASSP 2019accepted

Parkinson's disease (PD) and Amyotrophic Lateral Sclerosis (ALS) are progressive neurodegenerative diseases which, among other symptoms, cause dysarthria of speech. To assist the clinical diagnosis and treatment of neurological diseases, several studies have addressed the characterization and classi…

Cited by 0SourceScholar
2016

Exploiting low-dimensional structures to enhance DNN based acoustic modeling in speech recognition

ICASSP 2016accepted

We propose to model the acoustic space of deep neural network (DNN) class-conditional posterior probabilities as a union of low-dimensional subspaces. To that end, the training posteriors are used for dictionary learning and sparse coding. Sparse representation of the test posteriors using this dict…

Cited by 0SourceScholar
2016

System fusion and speaker linking for longitudinal diarization of TV shows

ICASSP 2016accepted

Performing speaker diarization while uniquely identifying the speakers in a collection of audio recordings is a challenging task. Based on our previous work on speaker diarization and linking, we developed a system for diarizing longitudinal TV show data sets based on the fusion of speaker diarizati…

Cited by 0SourceScholar
2015

Combining SGMM speaker vectors and KL-HMM approach for speaker diarization

ICASSP 2015accepted

In this paper, a method to use SGMM speaker vectors for speaker diarization is introduced. The architecture of the Information Bottleneck (IB) based speaker diarization is utilized for this purpose. The audio for speaker diarization is split into short uniform segments. Speaker vectors are obtained…

Cited by 5SourceScholar
2015

Novel GCC-PHAT model in diffuse sound field for microphone array pairwise distance based calibration

ICASSP 2015accepted

We propose a novel formulation of the generalized cross correlation with phase transform (GCC-PHAT) for a pair of microphones in diffuse sound field. This formulation elucidates the links between the microphone distances and the GCC-PHAT output. Hence, it leads to a new model that enables estimation…

Cited by 0SourceScholar
2015

Objective speech intelligibility assessment through comparison of phoneme class conditional probability sequences

ICASSP 2015accepted

Assessment of speech intelligibility is important for the development of speech systems, such as telephony systems and text-to-speech (TTS) systems. Existing approaches to the automatic assessment of intelligibility in telephony typically compare a reference speech signal to a degraded copy, which r…

Cited by 0SourceScholar
2015

On application of non-negative matrix factorization for ad hoc microphone array calibration from incomplete noisy distances

ICASSP 2015accepted

We propose to use non-negative matrix factorization (NMF) to estimate the unknown pairwise distances and reconstruct a distance matrix for microphone array position calibration. We develop new multiplicative update rules for NMF with incomplete input matrix that take into account the symmetry of the…

Cited by 0SourceScholar
2015

Robust microphone placement for source localization from noisy distance measurements

ICASSP 2015accepted

We propose a novel algorithm to design an optimum array geometry for source localization inside an enclosure. We assume a square-law decay propagation model for the sound acquisition so that the additive noise on the measured source-microphone distances is proportional to the distances regardless of…

Cited by 0SourceScholar