2026
MultiBanana: A Challenging Benchmark for Multi-Reference Text-to-Image Generation
CVPR 2026
Recent text-to-image generation models have acquired the ability of multi-reference generation and editing; that is, to inherit the appearance of subjects from multiple reference images and re-render them in new contexts. However, existing benchmark datasets often focus on generation using a single