← Search

Jasmine Collins

5 accepted papers

2026

What Makes a Good Generated Image? Investigating Human and Multimodal LLM Image Preference Alignment

AAAI 2026technical

Automated evaluation of generative text-to-image models remains a challenging problem. Recent works have proposed using multimodal LLMs to judge the quality of images, but these works offer little insight into how multimodal LLMs make use of concepts relevant to humans, such as image style or compos

Cited by 0SourcePDFScholar
2024

CommonCanvas: Open Diffusion Models Trained on Creative-Commons Images

CVPR 2024poster

We train a set of open text-to-image (T2I) diffusion models on a dataset of curated Creative-Commons-licensed (CC) images which yields models that are competitive with Stable Diffusion 2 (SD2). This task presents two challenges: (1) high-resolution CC images lack the captions necessary to train T2I…

Cited by 30SourcePDFScholar
2022

ABO: Dataset and Benchmarks for Real-World 3D Object Understanding

CVPR 2022poster

We introduce Amazon Berkeley Objects (ABO), a new large-scale dataset designed to help bridge the gap between real and virtual 3D worlds. ABO contains product catalog images, metadata, and artist-created 3D models with complex geometries and physically-based materials that correspond to real, househ…

Cited by 225PDFcodeScholar