← Search

Maja Taseska

2 accepted papers

2021

How To Design a Three-Stage Architecture for Audio-Visual Active Speaker Detection in the Wild

ICCV 2021poster

Successful active speaker detection requires a three-stage pipeline: (i) audio-visual encoding for all speakers in the clip, (ii) inter-speaker relation modeling between a reference speaker and the background speakers within each frame, and (iii) temporal modeling for the reference speaker. Each sta…

Cited by 62PDFcodeScholar
2015

Minimum Bayes risk signal detection for speech enhancement based on a narrowband DOA model

ICASSP 2015accepted

A desired speech signal in hands-free communication systems is often degraded by background noise and interferers. Data-dependent spatial filters for desired speech extraction depend on the power spectral density (PSD) matrices of the desired and the undesired signals, which are commonly estimated r…

Cited by 0SourceScholar