cd /news/artificial-intelligence/epiworld-grounding-llm-policy-agents… · home › topics › artificial-intelligence › article
[ARTICLE · art-145176] src=arxiv.org ↗ pub= topic=artificial-intelligence verified=true sentiment=↑ positive

EpiWorld: Grounding LLM Policy Agents in Epidemiological World Models

EpiWorld, a closed-loop framework that grounds an LLM policy actor in a learned action-conditioned epidemiological world model and a tiered skill library of public-health protocols, reduced cumulative hospitalisation by up to 59% across retrospective COVID-19 and Influenza datasets and by an average of ~16% across six LLM backbones, according to the arXiv:2610.02744v1 paper. The world model achieved the best out-of-distribution Peak-MAE among all forecasting baselines, and the framework outperformed reinforcement-learning and optimal-control policy baselines. The system distills simulated future outcomes into reusable lessons while keeping protocol constraints fixed, allowing policy decisions to improve without sacrificing interpretability or controllability.

by read1 min views2 publishedOct 5, 2026

arXiv:2610.02744v1 Announce Type: new Abstract: Epidemic intervention policies are textual artefacts that human decision-makers interpret, justify, and revise through natural language, making large language models a natural candidate for epidemic policy reasoning. A naive LLM, however, lacks the epidemic dynamics needed to project intervention consequences, the quantitative surveillance signals required to assess severity, and the institutional constraints that define admissible actions. We present EpiWorld, a closed-loop framework that grounds an LLM policy actor in a learned action-conditioned epidemiological world model and a tiered skill library of public-health protocols, surveillance tools, and adaptive lessons accumulated through after-action analysis. Given a candidate intervention, the world model predicts regional epidemic evolution and enables fast counterfactual rollouts that provide feedback for policy selection and refinement. Outcomes of simulated futures are distilled into reusable lessons while protocol constraints remain fixed, allowing the decision process to improve without sacrificing interpretability or controllability. We evaluate both the world model and the end-to-end framework on retrospective COVID-19 and Influenza datasets: the world model achieves the best out-of-distribution Peak-MAE among all forecasting baselines, and the closed-loop framework reduces cumulative hospitalisation by up to 59% across datasets and by an average of ~16% across six LLM backbones, outperforming reinforcement-learning and optimal-control policy baselines.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @epiworld 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/epiworld-grounding-l…] indexed:0 read:1min 2026-10-05 · —