2025
OVFact: Measuring and Improving Open-Vocabulary Factuality for Long Caption Models
EMNLP 2025
Large vision-language models (VLMs) often struggle to generate long and factual captions. However, traditional measures for hallucination and factuality are not well suited for evaluating longer, more diverse captions and in settings where ground-truth human-annotated captions are unavailable. We in