← Search

Diogo Fernandes Costa Silva

1 accepted papers

2026

TAGARELA - A PORTUGUESE SPEECH DATASET FROM PODCASTS

ICASSP 2026poster

Despite significant advances in speech processing, Portuguese remains under-resourced due to the scarcity of public, large-scale, and high-quality datasets. To address this gap, we present a new dataset, named TAGARELA, composed of over 8,972 hours of podcast audio, specifically curated for training…

Cited by 0SourcePDFScholar