2024
BRIDGE: Bridging Gaps in Image Captioning Evaluation with Stronger Visual Cues
ECCV 2024poster
"Effectively aligning with human judgment when evaluating machine-generated image captions represents a complex yet intriguing challenge. Existing evaluation metrics like CIDEr or CLIP-Score fall short in this regard as they do not take into account the corresponding image or lack the capability of…