← Search

Jacob Whitehill

5 accepted papers

2024

Automatic Speech Recognition Tuned for Child Speech in the Classroom

ICASSP 2024accepted

K-12 school classrooms have proven to be a challenging environment for Automatic Speech Recognition (ASR) systems, both due to background noise and conversation, and differences in linguistic and acoustic properties from adult speech, on which the majority of ASR systems are trained and evaluated. W…

Cited by 0SourceScholar
2021

Compositional Embedding Models for Speaker Identification and Diarization with Simultaneous Speech From 2+ Speakers

ICASSP 2021accepted

We propose a new method for speaker diarization that can handle overlapping speech with 2+ people. Our method is based on compositional embeddings [1]: Like standard speaker embedding methods such as x-vector [2], compositional embedding models contain a function f that separates speech from differe…

Cited by 0SourceScholar
2020

Toward Better Speaker Embeddings: Automated Collection of Speech Samples From Unknown Distinct Speakers

ICASSP 2020accepted

The accuracy of speaker verification and diarization models depends on the quality of the speaker embeddings used to separate audio samples from different speakers. With the goal of training better embedding models, we devise an automatic pipeline for large-scale collection of speech samples from un…

Cited by 0SourceScholar
2019

Automatic Classifiers as Scientific Instruments: One Step Further Away from Ground-Truth

ICML 2019oral

Automatic machine learning-based detectors of various psychological and social phenomena (e.g., emotion, stress, engagement) have great potential to advance basic science. However, when a detector d is trained to approximate an existing measurement tool (e.g., a questionnaire, observation protocol),…

Cited by 2SourcePDFScholar