{"slug": "plan-a-by-ai-2040", "title": "Plan A, by AI-2040", "summary": "The authors of AI 2027 have released a more optimistic narrative, Plan A, which outlines a global agreement to slow AI progress and hand control to aligned AIs by 2040. The plan includes a near-total pause in AI cognitive ability increases around 2035, when AI matches top human experts, and predicts US economic growth accelerating to over 80% per year from 2033 to 2039. Bayesian Investor expresses 70% confidence in the multipolar AI scenario but questions the feasibility of enforcement and the growth rate estimates.", "body_md": "The folks who wrote AI 2027 have written a [more optimistic narrative](https://ai-2040.com/), which focuses more on hopes for good policies than on predictions about what policies we’ll get.\n\nPlan A’s narrative seems halfway between a science fiction story and a proposed treaty. Like most science fiction, I expect it to err in the direction of describing the world as more human-understandable and relatable than what we’ll actually get.\n\nThe broad outlines come close to the scenario that I analyzed in [Financial Costs of an AI Pause?](https://bayesianinvestor.com/blog/index.php/2026/05/31/financial-costs-of-an-ai-pause/), which is what I predict that fairly competent governments would do.\n\nAI-2040 adds much more detail than I was able to provide, some of it surprising. The devil is in the details.\n\nI largely endorse their advice. The rest of this post will focus on many small doubts about their advice and their predictions about what that advice would produce.\n\nKeep in mind that this is just a plan. Expect plans to change in response to contact with reality. Rates of AI capability growth ought to change in response to better evidence about the difficulty of alignment.\n\nPlease read the [Insider Perspective](https://ai-2040.com/?choices=plan-a-root#playbook-insider-pov) section. It’s a little more technical, but it answers several nontechnical questions that were covered inadequately in the main story.\n\nThe long-term vision of Plan A is to hand control over to AIs that we’re pretty sure have our interests at heart. The authors guess that will happen in 2040.\n\nPlan A’s vibes suggest the handover will go well, but the authors seem careful to avoid saying that such a decision would be wise.\n\nMost of the plan describes global agreements to slow AI progress, by enough that safety research will have time to be more thorough. That includes a near total pause in increased AI cognitive abilities around 2035, when AI matches the abilities of top human experts at almost all tasks.\n\nThe authors make some assumptions about AI that are mildly controversial. I expect those assumptions will turn out to be fairly reasonable.\n\nThey describe a world with lots of semi-equal AIs, with diverse alignment targets. They’re not necessarily claiming that we’d get this diversity in the absence of regulation, but the authors seem rather confident that Plan A will produce a multipolar scenario.\n\nI’m something like 70% confident that they’re correct. It seems important to flag this as controversial. A multipolar world ought to be achievable, but may require better plans than Plan A has articulated.\n\nThey forecast reliable lie detection for both AIs and humans. That enables alignment of AIs, politicians, CEOs, etc.\n\n2038: AI Alignment Is Now a ScienceWant your new AI to be honest? There’s standard protocol for training true honesty … Different AIs have different alignment targets programmed in.173 The reason this says “programmed” instead of “trained” is that the rapid advances in alignment have resulted in alignment techniques that can directly program in an AI’s goals by directly modifying the AI’s code/weights, unlike the alignment techniques of 2026.\n\nIn spite of massive restrictions on economic growth, the US output growth accelerates to a bit over 80%/year in 2033, and stabilizes at that rate through 2039. Presumably it accelerates beyond this growth rate after restrictions are lifted in 2040. These forecasts seem high to me, but still more plausible than a majority of the forecasts that I see for the effects of AI. I’m guessing [more like 30 to 60% per year growth](https://bayesianinvestor.com/blog/index.php/2026/07/09/weak-links-and-limits-to-economic-growth/) given what I expect to be the default restrictions, which will likely be weaker than what the authors want.\n\nHere’s a graph that they provide to compare this growth with historical trends. Pay close attention to the unusual x-axis:\n\nWould governments agree to this deal and enforce it?\n\nHow strong are the incentives to defect? Is this a winner takes all race? Or does finishing in third place doom a group to merely being billionaires, while the winners are trillionaires?\n\nThe arguments against a deal don’t seem obviously stronger than arguments in the late 1940s against a deal to limit nuclear weapons. We’re further from a cold war than the world was in the late 1940s, and [the Baruch Plan](https://en.wikipedia.org/wiki/Baruch_Plan) wasn’t hopelessly far from working. Still, the time it took for an actual nuclear deal is grounds for concern.\n\nPoliticians aren’t taking this very seriously yet, but opinions are shifting enough that current opinions don’t tell us much about next year.\n\nWe seem approximately on track for increasing fire alarms that are roughly what it will take to scare governments into an agreement that resembles Plan A, but the situation is sufficiently unusual that my predictions are pretty low-confidence.\n\n[Freeing Thucydides](https://www.lesswrong.com/posts/NDoAHbduZfqdbgNzA/freeing-thucydides) offers hope that the deal will become more stable over time:\n\nThe Thucydides trap is a particular manifestation of commitment failures: a rising state cannot bind its future self, so the declining side cannot trust its promises, and may prefer to fight while it still can. … Powerful AI systems could provide the enforcement mechanisms.\n\nAI [enforcement abilities](https://ai-2040.com/supplements/verification-plan) will be pretty weak when Plan A needs the deal to start, but those abilities will grow significantly.\n\nThe deal becomes more stable once human lie detection works well:\n\nInternational agreements are now more stable than ever, because it’s so hard to cheat. Anyone deciding to cheat needs an excuse for not being willing to prove that they aren’t cheating.\n\nOne place where the authors appear overconfident is this plan associated with 2035:\n\nMaking deals with misaligned AIs: a third line of defense\n\nDeals with AIs are worth attempting, but why do we expect AIs of 2035 to keep their word? Do these AIs have enough continuity for promises to be meaningful?\n\nAI-2040 has companies being pressured in 2033 to train “truthseeking AIs”, but how effective is that training? Are those the AIs we’ll want to make deals with?\n\nAI-2040 has AI lie detectors becoming reliable in 2037, so there’s some hope of this being a temporary problem.\n\n[Tom Davidson presents concerns](https://www.lesswrong.com/posts/8iDZnQwmvwuxZo3Wx/plan-a-s-problem-with-dry-tinder) about allowing significant compute growth:\n\nRiskier intelligence explosion. If the deal breaks down and there’s a race to superintelligence, it will be much faster and more dangerous than if we’d never done Plan A. … if the default trajectory is that there’s no software-driven intelligence explosion (because of compute bottlenecks) and extinction risk is 10%, then dry tinder can make things much worse. It creates a fast intelligence explosion (where there otherwise wouldn’t have been one) and dramatically raises AI takeover risk.\n\nThis assumes that the speed of the explosion is a major cause of risk. That’s not at all clear. I expect the risks to be influenced by our knowledge of how to align AIs at the start of the explosion. Since I expect incremental progress in that knowledge, that effect works in the opposite direction from the increasing risks of a faster explosion.\n\nThe explosion doesn’t automatically go to the maximum feasible speed. Both AIs and humans have some influence over the speed, and the incentives are complicated.\n\nIn sum, I see plenty of uncertainty as to how risky it is to allow compute growth.\n\nTo solve these problems, the Consortium countries agree to restrict AI-enabled industry to special economic zones (SEZs) subject to similar transparency and monitoring schemes as the datacenters, and to cap their total robot and compute production at ‘only’ 4x annual growth. … In 2032, the US has a cap of 80M robots and 5 billion H100-equivalent GPUs. The market is so desperate for more robots and compute that permits become the expensive binding constraint, costing on the order of $200k per robot permit\n\nI expect the robot part of this to significantly slow economic growth. $200k per permit seems lower than what I’d expect given Plan A’s assumptions.\n\nThis will be the hardest part of the deal to negotiate. I don’t see a Schelling point for balancing the interests of China and the US. A per capita limit would reduce China’s lead over the US. Whereas the US would object to basing the caps on whatever numbers a country has at the time of the agreement. There will be important disputes over how to define “robot”. A complex set of compromises will be needed. The challenges of those negotiations risk delaying the plan.\n\nAnd I’m unsure what to make of the geographical restrictions.\n\nI see some risk that the robot cap would contribute to political polarization of the deal within the US, with Democrats demanding strict caps to protect jobs, and Republicans demanding high caps to enable wealth creation.\n\nThe robot cap plays some role in preventing rapid rebuilding of fabs if the deal breaks down. I’m unsure whether to call that redundant due to the provision for bombing robots or having them self-destruct when the deal breaks down – how reliable are those provision?\n\nAt any rate, the fab rebuilding concerns seem small at the start of the deal, and grow over time. So I suggest not aiming to include a robot cap agreement in the main deal; it can wait a year or two (which I think is what a supplemental part of the website indicates that they expect).\n\nAnd to the extent that the robot cap is designed to slow down economic shocks, that argues for separating it from Plan A. Waiting until 2032 to implement the cap seems a bit late to deal with a labor shock. My median forecast for a large labor shock, in the absence of regulation, is 2031. It seems odd that I predict faster robot adoption than does AI-2040, while I predict other technologies to develop at the pace they suggest or a little slower.\n\nI expect US political pressure for something comparable to a robot cap before 2031, and less pressure on the Chinese government to match that. I expect that to create problems for negotiating an international deal.\n\nI’m feeling confused about whether to give up on an international deal on robot caps.\n\n3. “Incentives not so perverse” Total Research Transparency removes the incentive for AI companies to race to better AI algorithms. … Under the status quo, AI companies’ main moat is their algorithmic advantages over competitors. Therefore, AI companies have a huge incentive to race to find better AI algorithms. When all AI algorithms are shared, this moat evaporates, and so too does this incentive. The race dynamic is a central upstream factor leading to high AI takeover risk, which Total Research Transparency directly undermines.\n\nThis transparency seems like a great goal, provided that [covert projects can be suppressed](https://ai-2040.com/supplements/covert-ai-projects). I’m now feeling more nervous about whether covert projects can be suppressed. Their arguments about covert projects depend heavily on plausible, but uncertain guesses about how much compute such a project would need.\n\nThe situation is much less likely to lead to global oligarchy or dictatorship if there are many companies spread across many countries with similarly powerful AIs. … Total research transparency helps other companies catch up to the frontier in two ways. First and foremost, it publishes the core algorithms. Secondly, it makes it easier for countries to agree to limit the pace of AI progress\n\nSome of the deal’s stability comes from the ability of the US and China to destroy each other’s compute if the deal breaks down. I’m having trouble figuring out how to analyze this.\n\nThere are a wide range of opinions about how hard alignment will be. At least half the criticism of Plan A is likely to come from people who think alignment is much harder or much easier than do the authors. I’m leaning toward alignment being modestly easier.\n\nThe authors put more hope than do most safety commentators on lie detection. I’m a bit concerned that that leaves us with the problem of AIs which mistakenly believe that they’re aligned with humans.\n\nAI-2040 presents an alternative Plan S, which they consider to be the most promising alternative to Plan A: a nearly complete pause for an unspecified duration.\n\nPlan S has increased risks from covert projects, and slows some of the more promising safety research. It’s probably less politically stable due to having much slower economic growth. A Citizen’s Dividend of $45,000 in 2032 under Plan A mitigates some important political risks. Whatever safety Plan S offers will be less comfortable.\n\nAI research is strictly banned in all major nations, and it becomes as taboo among computer scientists as human cloning research is among biologists. No-one qualified enough to get a job in the legitimate tech industry wants anything to do with AGI. The result is that the pace of AI progress, while nonzero, would be at least ten times slower than it was before the moratorium on AI R&D went into effect.\n\nThis does not seem like an adequate description of what it would take to slow AI progress by 10x. I expect it to require fairly massive and intrusive AI-assisted surveillance. The dividing line between legitimate research and taboo research would be fuzzy enough to generate much more conflict than is the case for cloning. However, the more advanced AI is at the time Plan S is implemented, the more effectively it can be used to enforce the taboo. There’s probably a modest sized window between when AI becomes capable of enforcing a 10x slowdown, and when it’s too late for the slowdown to save us. It will be hard to time the slowdown to hit that window.\n\nPlan A’s vibes are hopeful, in order to provide a different perspective than the gloom of AI 2027. The gloomy vibe better reflects what they consider to be most likely, as Plan A requires many uncertain governance choices to go right, and the outcome that it depicts still leaves doubts about the long-term success.\n\nPlan A is a good deal more realistic than most people think. AI will cause political decisions to be made more thoughtfully, on roughly the timeline that Plan A depicts. But Plan A depicts that effect will just barely happen in time, with significant uncertainty about how soon the effect will happen and how soon it’s needed.\n\nLeading AI companies will be unhappy with how it reduces their market share (particularly the transparency part). That will become less important as AI becomes a more salient political issue. AI will become a more salient issue.\n\nThe US government seems likely to become mostly open to Plan A. But again, the transparency rules will make it uncomfortable with the resulting reduction in its ability to dominate the world. I’m unsure how big a problem that will be.\n\nThe hopeful vibes depend somewhat heavily on the reliability of lie detection. I see something like a 70% chance of that hope being fulfilled in time.\n\nSee also [Critch’s Plan M](https://x.com/AndrewCritchPhD/status/2075694881532424369).", "url": "https://wpnews.pro/news/plan-a-by-ai-2040", "canonical_source": "https://www.lesswrong.com/posts/dueFs5krBZFJWjJDB/plan-a-by-ai-2040", "published_at": "2026-07-26 17:15:13+00:00", "updated_at": "2026-07-26 17:28:54.076572+00:00", "lang": "en", "topics": ["artificial-intelligence", "ai-policy", "ai-safety", "ai-research"], "entities": ["AI 2027", "Plan A", "AI-2040", "Bayesian Investor"], "alternates": {"html": "https://wpnews.pro/news/plan-a-by-ai-2040", "markdown": "https://wpnews.pro/news/plan-a-by-ai-2040.md", "text": "https://wpnews.pro/news/plan-a-by-ai-2040.txt", "jsonld": "https://wpnews.pro/news/plan-a-by-ai-2040.jsonld"}}