2024
BenchLMM: Benchmarking Cross-style Visual Capability of Large Multimodal Models
ECCV 2024poster
"Large Multimodal Models (LMMs) such as GPT-4V and LLaVA have shown remarkable capabilities in visual reasoning on data in common image styles. However, their robustness against diverse style shifts, crucial for practical applications, remains largely unexplored. In this paper, we propose a new benc…