00:00
2026-09-17
rocm.blogs.amd.com
ai-infrastructure
Implementing a High-Performance Custom Diffusion Attention Kernel with FlyDSL
AMD published a step-by-step FlyDSL workflow for implementing a custom diffusion attention kernel, targeting vLLM's TiDAR (Think in Diffusion, Talk in Autoregression) mode with paged KV caches and scr…