Unified Multimodal Autoregressive Modeling with Shared Context—Visual Tokenizer is Key to Unification
Unified Multimodal Modeling aims to integrate visual understanding and generation within a single system. However, existing approaches typically rely on two disparate visual tokenizers, which splits the representation space and hinder truly unified modeling. We propose UniAR, a unified autoregressiv…