← Search

Arkhipkin Vladimir

1 accepted papers

2024

Kandinsky 3: Text-to-Image Synthesis for Multifunctional Generative Framework

EMNLP 2024system demonstrations

Text-to-image (T2I) diffusion models are popular for introducing image manipulation methods, such as editing, image fusion, inpainting, etc. At the same time, image-to-video (I2V) and text-to-video (T2V) models are also built on top of T2I models. We present Kandinsky 3, a novel T2I model based on l…