← Search

Sharon Goldwater

12 accepted papers

2025

The Cross-linguistic Role of Animacy in Grammar Structures

ACL 2025long

Animacy is a semantic feature of nominals and follows a hierarchy: personal pronouns > human > animate > inanimate. In several languages, animacy imposes hard constraints on grammar. While it has been argued that these constraints may emerge from universal soft tendencies, it has been difficult to p…

2024

Estimating the Level of Dialectness Predicts Inter-annotator Agreement in Multi-dialect Arabic Datasets

ACL 2024short

On annotating multi-dialect Arabic datasets, it is common to randomly assign the samples across a pool of native Arabic speakers. Recent analyses recommended routing dialectal samples to native speakers of their respective dialects to build higher-quality datasets. However, automatically identifying…

2023

Analyzing Acoustic Word Embeddings from Pre-Trained Self-Supervised Speech Models

ICASSP 2023accepted

Given the strong results of self-supervised models on various tasks, there have been surprisingly few studies exploring self-supervised representations for acoustic word embeddings (AWE), fixed-dimensional vectors representing variable-length spoken word segments. In this work, we study several pre-…

Cited by 0SourceScholar
2021

[RETRACTED] Prosodic segmentation for parsing spoken dialogue

ACL 2021long

Parsing spoken dialogue poses unique difficulties, including disfluencies and unmarked boundaries between sentence-like units. Previous work has shown that prosody can help with parsing disfluent speech (Tran et al. 2018), but has assumed that the input to the parser is already segmented into senten…

2020

Analyzing ASR Pretraining for Low-Resource Speech-to-Text Translation

ICASSP 2020accepted

Previous work has shown that for low-resource source languages, automatic speech-to-text translation (AST) can be improved by pre-training an end-to-end model on automatic speech recognition (ASR) data from a high-resource language. However, it is not clear what factors - e.g., language relatedness…

Cited by 0SourceScholar
2020

Cross-Lingual Topic Prediction For Speech Using Translations

ICASSP 2020accepted

Given a large amount of unannotated speech in a low-resource language, can we classify the speech utterances by topicƒ We consider this question in the setting where a small amount of speech in the low-resource language is paired with text translations in a high-resource language. We develop an effe…

Cited by 0SourceScholar
2020

Multilingual Acoustic Word Embedding Models for Processing Zero-resource Languages

ICASSP 2020accepted

Acoustic word embeddings are fixed-dimensional representations of variable-length speech segments. In settings where unlabelled speech is the only available resource, such embeddings can be used in "zero-resource" speech search, indexing and discovery systems. Here we propose to train a single super…

Cited by 0SourceScholar
2017

Weakly supervised spoken term discovery using cross-lingual side information

ICASSP 2017accepted

Recent work on unsupervised term discovery (UTD) aims to identify and cluster repeated word-like units from audio alone. These systems are promising for some very low-resource languages where transcribed audio is unavailable, or where no written form of the language exists. However, in some cases it…

Cited by 0SourceScholar
2015

Unsupervised neural network based feature extraction using weak top-down constraints

ICASSP 2015accepted

Deep neural networks (DNNs) have become a standard component in supervised ASR, used in both data-driven feature extraction and acoustic modelling. Supervision is typically obtained from a forced alignment that provides phone class targets, requiring transcriptions and pronunciations. We propose a n…

Cited by 0SourceScholar