2025
Multi-modal Streaming ASR in Cross-talk Scenario for Smart Glasses
ICASSP 2025accepted
In the MMCSG task of the CHiME-8 Challenge, achieving real-time speaker-attributed transcriptions with limited multi-modal data presents significant challenges. To cope with the problem, we propose a novel ASR framework that leverages both audio-only and multi-modal inputs in a streaming fashion. Fo…