← Search

Chaerin Kong

3 accepted papers

2023

Unifying Vision-Language Representation Space with Single-Tower Transformer

AAAI 2023technical

Contrastive learning is a form of distance learning that aims to learn invariant features from two related representations. In this work, we explore the hypothesis that an image and caption can be regarded as two different views of the underlying mutual information, and train a model to learn a unif…

Cited by 20SourcePDFScholar
2022

Few-Shot Image Generation with Mixup-Based Distance Learning

ECCV 2022poster

"Producing diverse and realistic images with generative models such as GANs typically requires large scale training with vast amount of images. GANs trained with limited data can easily memorize few training samples and display undesirable properties like ""stairlike"" latent space where interpolati…