← Search

Lisen Dai

2 accepted papers

2025

CLIPErase: Efficient Unlearning of Visual-Textual Associations in CLIP

ACL 2025long

Machine unlearning (MU) has gained significant attention as a means to remove the influence of specific data from a trained model without requiring full retraining. While progress has been made in unimodal domains like text and image classification, unlearning in multimodal models remains relatively…

Cited by 0SourcePDFScholar
2024

SaSR-Net: Source-Aware Semantic Representation Network for Enhancing Audio-Visual Question Answering

EMNLP 2024finding

Audio-Visual Question Answering (AVQA) is a challenging task that involves answering questions based on both auditory and visual information in videos. A significant challenge is interpreting complex multi-modal scenes, which include both visual objects and sound sources, and connecting them to the…

Cited by 0SourcePDFScholar