2022
TaiSu: A 166M Large-scale High-Quality Dataset for Chinese Vision-Language Pre-training
NeurIPS 2022accept
Vision-Language Pre-training (VLP) has been shown to be an efficient method to improve the performance of models on different vision-and-language downstream tasks. Substantial studies have shown that neural networks may be able to learn some general rules about language and visual concepts from a la…