← Search

Anup Singh

2 accepted papers

2025

BEST-STD: Bidirectional Mamba-Enhanced Speech Tokenization for Spoken Term Detection

ICASSP 2025accepted

Query-by-example spoken term detection (QbE-STD) is often hindered by reliance on frame-level features and the computationally intensive DTW-based template matching, limiting its practicality. To address these challenges, we propose a novel approach that encodes speech into discrete, speaker-agnosti…

Cited by 0SourceScholar
2023

Simultaneously Learning Robust Audio Embeddings and Balanced Hash Codes for Query-by-Example

ICASSP 2023accepted

Audio fingerprinting systems must efficiently and robustly identify query snippets in an extensive database. To this end, state-of-the-art systems use deep learning to generate compact audio fingerprints. These systems deploy indexing methods, which quantize fingerprints to hash codes in an unsuperv…

Cited by 0SourceScholar