2024
Prometheus-Vision: Vision-Language Model as a Judge for Fine-Grained Evaluation
ACL 2024findings
Assessing long-form responses generated by Vision-Language Models (VLMs) is challenging. It not only requires checking whether the VLM follows the given instruction but also verifying whether the text output is properly grounded on the given image. Inspired by the recent approach of evaluating LMs w…