2025
LEDiT: Your Length-Extrapolatable Diffusion Transformer without Positional Encoding
NeurIPS 2025poster
Diffusion transformers (DiTs) struggle to generate images at resolutions higher than their training resolutions. The primary obstacle is that the explicit positional encodings (PE), such as RoPE, need extrapolating to unseen positions which degrades performance when the inference resolution differs…