cd /news/artificial-intelligence/dead-text-or-binding-clause-measurin… · home topics artificial-intelligence article
[ARTICLE · art-96280] src=arxiv.org ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

Dead text or binding clause? Measuring and restoring constraint influence in black-box LLM dialogues

A new arXiv preprint (2608.12599v1) introduces a method to measure and restore constraint influence in black-box LLM dialogues, finding that behavioral relapse—models continuing to enact revoked constraints—increases with constraint load at an 8B operating point, while stronger models remain at floor. The method, using a contract ledger, sequential ablation probe, and repair ladder, significantly reduces relapse against a no-ledger baseline (95% CI, p < 0.05), with a one-sentence tombstone note recovering about a third of the effect. The probe predicts relapse before delivery (AUROC 0.85), and the approach operates at 1.2x delivery overhead and 0.3% of API compute.

read1 min views1 publishedAug 14, 2026

arXiv:2608.12599v1 Announce Type: new Abstract: Multi-turn dialogues let users revoke constraints as easily as impose them, but revocation does not reliably take effect: models keep enacting withdrawn requirements (occasionally beneath comments asserting their removal), a failure we call \emph{behavioral relapse}, or revocation inertia. No existing instrument measures this influence per clause, predicts it before delivery, or repairs it under matched budgets. \sysname{} closes the three gaps through the model API alone: a contract ledger pairs every constraint with an executable checker, records revocations as tombstones, and compiles the net constraint state ahead of time into a single specification; a sequential ablation probe measures per-clause adherence and incremental behavioral effect; a repair ladder operates under token- and attempt-matched budgets. On \dataname{} (\NTasks{} HumanEval tasks, \NClauses{} verified checkers), relapse at an 8B operating point climbs from \ScaleDelayedMTwo{} to \ScaleDelayedMEight{} as constraint load grows, while stronger models sit at floor. Under matched checkers, model, and budget, ahead-of-time compilation significantly reduces relapse against a no-ledger verifier-retry baseline (\RestoreDiff{}, 95% CI \RestoreDiffCI{}, $p$ \RestoreDiffP{}); adaptive ladder interventions stacked on top add no detectable gain (95% confidence excludes gains $\geq$ \LadderExcludedGain{}). The probe predicts relapse before delivery (AUROC \AurocPrimary{}); a one-sentence tombstone note recovers about a third of the compilation effect and survives a placebo control. At \CostDeliveryFactor{} delivery overhead and \CostTotalHedged{} of API compute for every result, revocation failure becomes a measurable, predictable, and repairable property of dialogue state rather than an invisible one.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @arxiv 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/dead-text-or-binding…] indexed:0 read:1min 2026-08-14 ·