← Search

Abhijeet Sangwan

4 accepted papers

2019

Transfer Learning Using Raw Waveform Sincnet for Robust Speaker Diarization

ICASSP 2019accepted

Speaker diarization tells who spoke and when? in an audio stream. SincNet is a recently developed novel convolutional neural network (CNN) architecture where the first layer consists of parameterized sinc filters. Unlike conventional CNNs, SincNet take raw speech waveform as input. This paper levera…

Cited by 0SourceScholar
2018

Robust Feature Clustering for Unsupervised Speech Activity Detection

ICASSP 2018accepted

In certain applications such as zero-resource speech processing or very-low resource speech-language systems, it might not be feasible to collect speech activity detection (SAD) annotations. However, the state-of-the-art supervised SAD techniques based on neural networks or other machine learning me…

Cited by 0SourceScholar
2015

Prof-Life-Log: Analysis and classification of activities in daily audio streams

ICASSP 2015accepted

A new method to analyze and classify daily activities in personal audio recordings (PARs) is presented. The method employs speech activity detection (SAD) and speaker diarization systems to provide high level semantic segmentation of the audio file. Subsequently, a number of audio, speech and lexica…

Cited by 19SourceScholar
2015

Robust overlapped speech detection and its application in word-count estimation for Prof-Life-Log data

ICASSP 2015accepted

The ability to estimate the number of words spoken by an individual over a certain period of time is valuable in second language acquisition, healthcare, and assessing language development. However, establishing a robust automatic framework to achieve high accuracy is non-trivial in realistic/natura…

Cited by 0SourceScholar