2025
CogCM: Cognition-Inspired Contextual Modeling for Audio-Visual Speech Enhancement
ICCV 2025poster
Audio-Visual Speech Enhancement (AVSE) leverages both audio and visual information to improve speech quality. Despite noisy real-world conditions, humans are generally able to perceive and interpret corrupted speech segments as clear. Researches in cognitive science have shown how the brain merges a…