2025
Revisiting text-to-image evaluation with Gecko: on metrics, prompts, and human rating
ICLR 2025spotlight
While text-to-image (T2I) generative models have become ubiquitous, they do not necessarily generate images that align with a given prompt. While many metrics and benchmarks have been proposed to evaluate T2I models and alignment metrics, the impact of the evaluation components (prompt sets, human…