CROPE: Evaluating In-Context Adaptation of Vision and Language Models to Culture-Specific Concepts
As Vision and Language models (VLMs) become accessible across the globe, it is important that they demonstrate cultural knowledge. In his paper, we introduce CROPE, a visual question answering benchmark designed to probe the knowledge of culture-specific concepts and evaluate the capacity for cultur…