04:00
2026-08-04
arxiv.org
artificial-intelligence
DLLM-TTS: Block Discrete Diffusion Language Model for Text-to-Speech Synthesis
Researchers introduced DLLM-TTS, a text-to-speech framework using block discrete diffusion over X-Codec2 neural audio codec tokens, achieving a real-time factor of 0.15 with a 0.6B-parameter model traβ¦