2024
Explaining CLIP's Performance Disparities on Data from Blind/Low Vision Users
CVPR 2024poster
Large multi-modal models (LMMs) hold the potential to usher in a new era of automated visual assistance for people who are blind or low vision (BLV). Yet these models have not been systematically evaluated on data captured by BLV users. We address this by empirically assessing CLIP a widely-used LMM…