← Search

Rizhao Cai*

1 accepted papers

2024

BenchLMM: Benchmarking Cross-style Visual Capability of Large Multimodal Models

ECCV 2024poster

"Large Multimodal Models (LMMs) such as GPT-4V and LLaVA have shown remarkable capabilities in visual reasoning on data in common image styles. However, their robustness against diverse style shifts, crucial for practical applications, remains largely unexplored. In this paper, we propose a new benc…