← Search

Hira Dhamyal

4 accepted papers

2024

On the Evaluation of Speech Foundation Models for Spoken Language Understanding

ACL 2024findings

The Spoken Language Understanding Evaluation (SLUE) suite of benchmark tasks was recently introduced to address the need for openresources and benchmarking of complex spoken language understanding (SLU) tasks, including both classification and sequence generation tasks, on natural speech. The benchm…

Cited by 6SourcePDFScholar
2024

Prompting Audios Using Acoustic Properties for Emotion Representation

ICASSP 2024accepted

Emotions lie on a continuum, but current models treat emotions as a finite valued discrete variable. This representation does not capture the diversity in the expression of emotion. To better represent emotions we propose the use of natural language descriptions (or prompts). In this work, we addres…

Cited by 0SourceScholar
2024

R-BASS : Relevance-aided Block-wise Adaptation for Speech Summarization

NAACL 2024findings

End-to-end speech summarization on long recordings is challenging because of the high computational cost. Block-wise Adaptation for Speech Summarization (BASS) summarizes arbitrarily long sequences by sequentially processing abutting chunks of audio. Despite the benefits of BASS, it has higher compu…

Cited by 0SourcePDFScholar
2024

Speech vs. Transcript: Does It Matter for Human Annotators in Speech Summarization?

ACL 2024long

Reference summaries for abstractive speech summarization require human annotation, which can be performed by listening to an audio recording or by reading textual transcripts of the recording. In this paper, we examine whether summaries based on annotators listening to the recordings differ from tho…