← Search

Christopher Simic

3 accepted papers

2025

Adapter-Based Multi-Agent AVSR Extension for Pre-Trained ASR Models

ICASSP 2025accepted

We present an approach to Audio-Visual Speech Recognition that builds on a pre-trained Whisper model. To infuse visual information into this audio-only model, we extend it with an AV fusion module and LoRa adapters, one of the most up-to-date adapter approaches. One advantage of adapter-based approa…

Cited by 0SourceScholar