← Search

Ramon Sanabria

6 accepted papers

2023

Analyzing Acoustic Word Embeddings from Pre-Trained Self-Supervised Speech Models

ICASSP 2023accepted

Given the strong results of self-supervised models on various tasks, there have been surprisingly few studies exploring self-supervised representations for acoustic word embeddings (AWE), fixed-dimensional vectors representing variable-length spoken word segments. In this work, we study several pre-…

Cited by 0SourceScholar
2023

The Edinburgh International Accents of English Corpus: Towards the Democratization of English ASR

ICASSP 2023accepted

English is the most widely spoken language in the world, used daily by millions of people as a first or second language in many different contexts. As a result, there are many varieties of English. Although the great many advances in English automatic speech recognition (ASR) over the past decades,…

Cited by 49SourceScholar
2019

Multimodal Grounding for Sequence-to-sequence Speech Recognition

ICASSP 2019accepted

Humans are capable of processing speech by making use of multiple sensory modalities. For example, the environment where a conversation takes place generally provides semantic and/or acoustic context that helps us to resolve ambiguities or to recall named entities. Motivated by this, there have been…

Cited by 0SourceScholar
2018

Sequence-Based Multi-Lingual Low Resource Speech Recognition

ICASSP 2018accepted

Techniques for multi-lingual and cross-lingual speech recognition can help in low resource scenarios, to bootstrap systems and enable analysis of new languages and domains. End-to-end approaches, in particular sequence-based techniques, are attractive because of their simplicity and elegance. While…

Cited by 0SourceScholar