2023
Open-vocabulary Object Segmentation with Diffusion Models
ICCV 2023poster
The goal of this paper is to extract the visual-language correspondence from a pre-trained text-to-image diffusion model, in the form of segmentation map, i.e., simultaneously generating images and segmentation masks for the corresponding visual entities described in the text prompt. We make the fol…