LUDI Framework Scales Diffusion Language Models with Per-Token Embeddings
September 30, 2026
The LUDI framework addresses scaling issues in uniform diffusion language models by introducing a less uniform loss and per-token time embeddings. The resulting LUDI-7B model achieves a 3-token-per-step speedup over autoregressive decoding while maintaining competitive performance.
HOW THIS AFFECTS YOU
●
builderThis offers a potential path to faster-than-AR inference for complex reasoning tasks.
●
researcherYou can explore non-autoregressive scaling using per-token corruption hints.