← Search

Tobias Cord-Landwehr

4 accepted papers

2025

Simultaneous Diarization and Separation of Meetings through the Integration of Statistical Mixture Models

ICASSP 2025accepted

We propose an approach for simultaneous diarization and separation of meeting data. It consists of a complex Angular Central Gaussian Mixture Model (cACGMM) for speech source separation, and a von-Mises-Fisher Mixture Model (vMFMM) for diarization in a joint statistical framework. Through the integr…

Cited by 0SourceScholar
2024

Geodesic Interpolation of Frame-Wise Speaker Embeddings for the Diarization of Meeting Scenarios

ICASSP 2024accepted

We propose a modified teacher-student training for the extraction of frame-wise speaker embeddings that allows for an effective diarization of meeting scenarios containing partially overlapping speech. To this end, a geodesic distance loss is used that enforces the embeddings computed from regions w…

Cited by 0SourceScholar
2023

Frame-Wise and Overlap-Robust Speaker Embeddings for Meeting Diarization

ICASSP 2023accepted

Using a Teacher-Student training approach we developed a speaker embedding extraction system that outputs embeddings at frame rate. Given this high temporal resolution and the fact that the student produces sensible speaker embeddings even for segments with speech overlap, the frame-wise embeddings…

Cited by 0SourceScholar
2021

Contrastive Predictive Coding Supported Factorized Variational Autoencoder For Unsupervised Learning Of Disentangled Speech Representations

ICASSP 2021accepted

In this work we address disentanglement of style and content in speech signals. We propose a fully convolutional variational autoencoder employing two encoders: a content encoder and a style encoder. To foster disentanglement, we propose adversarial contrastive predictive coding. This new disentangl…

Cited by 24SourceScholar