cd /news/large-language-models/plan-and-patch-diffusion-language-mo… · home › topics › large-language-models › article
[ARTICLE · art-148039] src=arxiv.org ↗ pub= topic=large-language-models verified=true sentiment=↑ positive

Plan-and-Patch: Diffusion Language Models for Agentic Planning

A diffusion language model (dLLM) planner achieved a 53.7% plan repair success rate versus 27.0% for an autoregressive (AR) planner on Natural Plan without task-specific training, according to the arXiv paper 2610.10786v1 introducing the Plan-and-Patch framework. Comparing DreamReasoner-8B (diffusion) and Qwen3-8B (AR), the paper reports that after task-specific training on ALFWorld and TextCraft the planners reached similar plan-generation success while diffusion cut mean plan-generation latency by 39-46%. Plan-and-Patch generates structured, program-like plans through parallel unmasking and repairs them by regenerating only selected regions while keeping the surrounding prefix and suffix fixed.

by read1 min views3 publishedOct 9, 2026

arXiv:2610.10786v1 Announce Type: new Abstract: Planning is increasingly important for long-horizon agents, where successful execution requires coordinating subgoals, tool use, and intermediate outcomes over many steps. Yet assumptions made during planning may be invalidated by the environment, tools may return unexpected results, or actions may fail. Effective agents must therefore not only generate plans, but also revise them. Such revisions often affect only part of a plan, leaving the preceding and subsequent structure intact. Rather than regenerate the entire plan and risk unnecessary changes, repair can regenerate the affected region conditioned on the preserved prefix and suffix. We introduce Plan-and-Patch, a plan-and-act framework in which a diffusion language model (dLLM) generates a structured, program-like plan through parallel unmasking and repairs it by filling in selected regions while keeping the surrounding steps fixed. We compare DreamReasoner-8B and Qwen3-8B as diffusion and autoregressive (AR) planners. On Natural Plan without task-specific training, diffusion (53.7%) achieves nearly twice the plan repair success rate of AR (27.0%). After task-specific training on agentic benchmarks, ALFWorld and TextCraft, the planners achieve similar observed success in plan generation, while diffusion reduces mean plan-generation latency by 39-46% relative to AR. Our results show that Plan-and-Patch provides a framework for faster plan generation and effective plan repair in long-horizon agents.

── more in #large-language-models 4 stories · sorted by recency
── more on @plan-and-patch 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/plan-and-patch-diffu…] indexed:0 read:1min 2026-10-09 · —