2025
FineGRAIN: Evaluating Failure Modes of Text-to-Image Models with Vision Language Model Judges
NeurIPS 2025spotlight
Text-to-image (T2I) models are capable of generating visually impressive images, yet they often fail to accurately capture specific attributes in user prompts, such as the correct number of objects with the specified colors. The diversity of such errors underscores the need for a hierarchical evalua…