← Search

Tetsugo Inada

1 accepted papers

2024

ConCon-Chi: Concept-Context Chimera Benchmark for Personalized Vision-Language Tasks

CVPR 2024poster

While recent Vision-Language (VL) models excel at open-vocabulary tasks it is unclear how to use them with specific or uncommon concepts. Personalized Text-to-Image Retrieval (TIR) or Generation (TIG) are recently introduced tasks that represent this challenge where the VL model has to learn a conce…