2026
Rainbow Padding: Mitigating Early Termination in Instruction-Tuned Diffusion LLMs
ICLR 2026poster
Diffusion large language models (dLLMs) have emerged as a promising alternative to autoregressive models, offering flexible generation orders and strong performance on complex reasoning tasks. However, instruction-tuned dLLMs exhibit a critical vulnerability we term \<eos\> overflow: as allocated s…