LLaDA-UI: Bringing Block-wise Diffusion to Vision-Language GUI Agents Researchers introduced LLaDA-UI, a block-wise diffusion vision-language model designed for GUI agents that must repeatedly perceive screen states and emit actions. The work applies diffusion large language models' block-parallel, arbitrary-order generation to latency-sensitive GUI agent tasks, aiming to improve decoding efficiency. No performance figures or benchmark results were disclosed in the available source material. Diffusion large language models dLLMs achieve high decoding efficiency through block-parallel, arbitrary-order generation, making them attractive for latency-sensitive applications. GUI agents represent a natural testbed for this paradigm, as they must repeatedly perceive screen states and emit st