← Search

Maris Basha

2 accepted papers

2026

VocSim A Training-free Benchmark for Zero-shot Content Identity in Single-source Audio

ICML 2026poster

General-purpose audio representations aim to map acoustically variable instances of the same event to nearby points, resolving content identity in a zero-shot setting. Unlike supervised classification benchmarks that measure adaptability via parameter updates, we introduce VocSim, a training-free be…

Cited by 0SourceScholar
2024

Positive Transfer of the Whisper Speech Transformer to Human and Animal Voice Activity Detection

ICASSP 2024accepted

This paper introduces WhisperSeg, utilizing the Whisper Transformer pre-trained for Automatic Speech Recognition (ASR) for human and animal Voice Activity Detection (VAD). Contrary to traditional methods that detect human voice or animal vocalizations from a short audio frame and rely on careful thr…

Cited by 0SourceScholar