← Search

Pooneh Mousavi

1 accepted papers

2025

What Are They Doing? Joint Audio-Speech Co-Reasoning

ICASSP 2025accepted

In audio and speech processing, tasks usually focus on either the audio or speech modality, even when both sounds and human speech are present in the same audio clip. Recent Auditory Large Language Models (ALLMs) have made it possible to process audio and speech simultaneously within a single model,…

Cited by 0SourceScholar