2025
VL-RewardBench: A Challenging Benchmark for Vision-Language Generative Reward Models
CVPR 2025highlight
Vision-language generative reward models (VL-GenRMs) play a crucial role in aligning and evaluating multimodal AI systems, yet their own evaluation remains under-explored. Current assessment methods primarily rely on AI-annotated preference labels from traditional VL tasks, which can introduce biase…