← Search

Jingwu Xiao

1 accepted papers

2024

Autoregressive Pre-Training on Pixels and Texts

EMNLP 2024main

The integration of visual and textual information represents a promising direction in the advancement of language models. In this paper, we explore the dual modality of language—both visual and textual—within an autoregressive framework, pre-trained on both document images and texts. Our method empl…