# Unlocking Lossless Speedups in LLMs via Discrete Diffusion

> Source: <https://aiflash.com/news/115430/>
> Published: 2026-09-08 08:00:25+00:00

Large Language Models (LLMs) owe much of their success to next-token prediction (NTP), but their autoregressive (AR) structure requires slow, sequential token generation. To overcome this bottleneck, we introduce diffusion-augmented LLMs, a new class of models that defines an AR model distribution w
