← Search

Tony Cheng Tong

1 accepted papers

2025

G-VEval: A Versatile Metric for Evaluating Image and Video Captions Using GPT-4o

AAAI 2025technical

Evaluation metric of visual captioning is important yet not thoroughly explored. Traditional metrics like BLEU, METEOR, CIDEr, and ROUGE often miss semantic depth, while trained metrics such as CLIP-Score, PAC-S, and Polos are limited in zero-shot scenarios. Advanced Language Model-based metrics als…