Some of these do seem like misrepresentations. For example, about your beliefs regarding to whom the economic benefits of AI will accrue, and also about the desirability of having a central planner.
Others seem mainly like disagreements about what the likely effect of proposed policies would be and/or qualitative judgments about those policies.
For example, it’s possible that mandating transparency on AI labs gives more room for policymakers to iterate, but it’s also a very big upfront change. It’s plausible someone could consider that change to bake in too much upfront, without misrepresenting the plan.
It’s also possible that the existing powers the government has are comparable to or greater than the various powers you propose, so that it’s no big deal to give them more powers. Others might disagree. This also seems more like an area of policy disagreement rather than misrepresentation.
I might have misunderstood which are misrepresentations and which are disagreements. Either way, I would be interested in some clarity as to which of these matters you vehemently disagree with Séb on and which you feel are misrepresentations.
Thanks for this comment!
I think the central disagreement is about whether AGI/ASI is real or not, and almost everything else is just downstream of that. The claims and policies made in AI 2040 only make sense in a world where you can have AIs that are much much smarter than humans, and that mere AGIs can cause explosive growth.
Given that Seb disagrees with the economics substantially, I would guess that he thinks that AIs will never reach the point where they can fully automate the robot and semiconductor supply chain; if we do get full automation then fast exponential growth seems to quickly fall out of reasonable modeling. Unfortunately I don't understand his views that well, so it's hard for me to say exactly where he thinks AI capabilities will cap out.
Re: Transparency and government powers, I think his post pretty badly misrepresented what we were saying, where I think someone who read his piece but not ours would be extremely misled about what sorts of regulations we were proposing. But I also imagine there's substantial policy disagreement there... so from my perspective it seems like both?
I will say that with government powers in particular, Plan A does involve governments into AI more than the status quo. Insofar as you think that'll be predictably bad, that is a genuine downside of Plan A. There's a sense in which any regulation is "central planning"... and yes, we do advocate for certain types of AI regulation. But overall in Plan A we aim for as market based and decentralized solutions: for example, we propose several cap and trade regimes, which seem to me the maximally libertarian way to go about this regulation, and so calling it "central planning" seems pretty misleading. Other regulation is more difficult to set up this way, such as limitations on algorithmic progress.
Thanks for this reply. I do agree that many people who are skeptical of various AI safety proposals don’t really believe in AGI or ASI, but I know that’s not everyone. For example, I am skeptical of AI 2040’s proposals but I very much believe in AGI and ASI, and that’s at the heart of some of the reasons for my concerns. I don’t know if that’s true in Séb’s case: is there a particular thing he wrote which makes you think he does not, or are you speculating/generalizing from experience with others?
Regarding the misrepresentations vs. disagreements, what I would really appreciate is understanding which are which in this post. I count 16 quotes from Séb listed under “False Representations”, but some of them seem more like disagreements to me. It would be helpful if you could clarify which you believe are actual misrepresentations and which you believe are disagreements.
I think it's true that Séb doesn't believe in AGI/ASI in the sense that I mean: see his fourth response here. Also see e.g. this model he posted which argues that comparative advantage implies postASI humans will still have jobs (which doesn't make sense if you think ASI can use the inputs that humans require much more efficiently than the humans do).
Seems very reasonable to be skeptical of many of AI 2040s proposals despite believing in AGI/ASI. I'd be curious to hear more (though perhaps you should post those on the main post, not here, if you feel interested :) ).
Re clarifing the quotes; reading them now I still think they are all misleading but to varying degrees and probably won't spend more time further getting into it because there's a lot on my plate right now, but if others think it'd be useful for me or someone else at AIFP to do another pass here trying to further clarify let us know (e.g. by reacting or commenting here).
I don't think belief in ASI necessarily implies belief in all of:
> (a) extremely fast diffusion and societal transformation, (b) a view that all profits accrue maximally to the labs, (c) that the prescriptions advanced in the essay (like expropriations and forced IP diffusion) have minimal impacts on said profits; (d) the claim that you get explosive GDP growth very soon, (e) that ‘de facto’ nationalisation and profit redistribution through UBI is an optimal response.
Or did you mean some other part of his fourth response?
I also don't think it requires you believe that no comparative advantage for humans post-ASI exists: positional goods are still a thing, a price premium for human-produced goods is very plausible (and positing that it won't happen is not a question of capabilities).
I think it's a common move to claim that people with a different conception of a post-ASI trajectory "don't believe in" ASI, but when digging into it usually they don't disagree on raw capabilities, just on what those raw capabilities imply the ASI would be able to do in the world.
(a) extremely fast diffusion and societal transformation, (b) a view that all profits accrue maximally to the labs, (c) that the prescriptions advanced in the essay (like expropriations and forced IP diffusion) have minimal impacts on said profits; (d) the claim that you get explosive GDP growth very soon, (e) that ‘de facto’ nationalisation and profit redistribution through UBI is an optimal response.
As I said in the post, this part is largely a strawman of our views. The part I mostly meant was "Nor do I think we will get real GDP growth of 50% in 2032" and "the model moves far too quickly from AIs being able to perform tasks to robots being reliable, legally deployable, organisationally integrated substitutes for almost all labour, and from there to a closed-loop reproduction of capital". Perhaps what he means is that there will be delays, but eventually we'll get 50% GDP growth?
I think it's a common move to claim that people with a different conception of a post-ASI trajectory "don't believe in" ASI, but when digging into it usually they don't disagree on raw capabilities, just on what those raw capabilities imply the ASI would be able to do in the world.
I tend to find the exact opposite is true. People love claiming that the real disagreement is real world bottlenecks, but usually its really differences in capability expectations. I think the AIs will be wildly superintelligent, with vast quantities operating at the equivalent of 100x or 1000x human speeds, also with a much higher qualitative intelligence due to things like knowing way more than any human can know, having much more experience than any human can every get, and having a physically much larger and more connected brain such that they are able to make discoveries and model things that are far too complicated for humans.
This is obviously an extreme milestone; we outline earlier milestones here: https://ai-rates-calculator.vercel.app/, for example. I would encourage people to make forecasts of when they think these specific milestones will be crossed (which might be "in hundreds of years or never"). If earlier than that, maybe its just a timelines / takeoff speed disagreement.
I kind of expect that many people would argue that this is the wrong ontology for thinking about AI capability progressions, and that actually, it's somehow going to be more diffuse/multipolar or something, in a way where thinking in these terms isn't useful for modeling the world, in which case I'd love to hear a better frame.
Still, I think that most people with very different views would tend to disagree that the notion of " superintelligent" AI that I outlined above would happen but will have small effects on the world. I think this view is extremely hard to make coherent. Once we've got AIs like that, the cognitive labour supply would become enourmous, such that almost all the cognitive labour happening would be AIs, not humans.
Totally appreciate you guys have a lot on your plate right now! At least two of the examples I mentioned in my top-level comment didn’t seem like misrepresentations to me (though two of them did), which makes me confused as to how much of this post is accurate if it really is saying that all of the quoted examples are misrepresentations (hence the reason for my comment requesting that clarification). I haven’t carefully dug into every claim myself but overall my sense is that this post very heavily blurs the line between vehement disagreement and actual misrepresentation such that I don’t feel there’s a clear takeaway for me about how much of each is going on in Séb’s critique.
After reflection, I decided to remove the two false representations that were the most ambiguous from the top level post. They were:
>A cadre of elites decides which research directions are permissible, caps global compute and robotics, and creates state-administered scarcity rents.
>It feels like the same plans as I’ve been hearing about for nearly a decade in the AI safety community, but filled with more details. In 2018 when I worked for the UK government, a prominent AI safety research organisation told me that “we need to solve the technical alignment problem, and then simply hand it to the UN to implement everywhere.” This was before the field invested in governance and politics; Plan A broadly similar, but with all sorts of mechanisms to fill in the gaps.
I think these are still wrong/misleading, but are less clear cut than the examples still in the post.
(We also edited the post to be generally be less combative, because I think it was too combative and regret that; see the italics at the top; you can see the diff here)
In the case of AI, training code and training data is far more auditable than the resulting model weights, which are (for now) an uninterpretable pile of numbers. So if you're a user who wants to confirm that your AI is working for you and not pursuing some covert agenda, total research transparency will do you a lot more good than mandatory open weights.
If I were trying to do this today and I had the choice, I would take the weights and not the data + algorithms. We've a reasonable collection of interp techniques that demand weights only, and a much smaller collection of techniques that demands data + training code only. I may be wrong, data + code is certainly helpful and both together is clearly better that either alone, but that's my best guess.
(assuming I can't actually go ahead and train a model on a significant fraction of that data)
Thanks! My intuition is the opposite, but I'd be curious to know which interp techniques you think are most valuable right now.
Though actually I realized that paragraph in the post is somewhat misleading (in a way which strengthens your point), sorry about that! Total research transparency gives you ~all the code, and some of the training data, but most of the training data is in the opaque database, and so can't be fully audited except via AIs, which might still get you a bunch of the benefit. (Read our detailed proposal here: https://ai-2040.com/supplements/transparency-plan. )
So this is a guess, but:
Anyway I would forcefully make the weak claim that high confidence that data + code > weights or the reverse is not justified.
I’m surprised by how unanimously people seem to think they can design a data audit that will be able to alert them if OLMo is reward hacking on their task
(edit: this opinion no longer seems so unanimous)
You miss the point. Suppose that Claude Bookshelf 7 and DeepSeekv7 have the same capabilities, and Claude Bookshelf 7 has the entire training run audited and DeepSeekv7 is open-sourced. Then DeepSeekv7 gets finetuned by terrorists, but safetyists fail to evaluate it for reasons similar to the last paragraph of METR's report on GPT-5.6 Sol: "As training and iteration continues, we need to ensure the models aren’t just learning to be more successful at evading the monitoring system. This is impossible to validate in a traditional pre-deployment evaluation paradigm, as it requires deep access to internal systems."
I did say "if I were trying to do this today". We might disagree about how relevant this is: I think "what I would prefer today" has a little more evidential weight than "plausible exploration of what might happen in the future" (though neither have all that much weight), perhaps you think it doesn't.
This response is too angry and generally not a good look. I don't have a position on the object level at the moment, but reading some of this response makes me think more favorably of Séb; a lot of what you say seems like disingenuous nitpicking and yelling "we didn't say that!!!" when you said something quite similar and it's clear Séb simply disagrees.
For example, you say:
Much of this is wrong. Specifically, we believe: (b) not all profits will accrue to labs,[1] (c) that Plan A would significantly decrease lab profits,[2] and (e) ‘de facto’ nationalization is not an optimal response.[3]
This is kind of ridiculous: you believe (b) not all profits accrue to the labs, because some accrue to NVIDIA and a tiny handful of other firms. Séb disagrees, clearly. But instead of addressing your disagreement productively you accuse him of lying. (c) You do believe that Plan A would not significantly decrease the relative profits, i.e. you think that even with the extreme slowdown, the AI labs eat the world. Séb disagrees. (e) I do think Plan A sounds like de facto nationalization, and I think it's disingenuous to imply that the government can already do more (it costs the government a large amount of political capital to try and they may also lose in court).
I'm disappointed by this response.
I thought both this piece and Séb's were unreasonable in multiple places. E.g. I thought this was tendentious:
As for our plan baking in too much and not leaving room for trial-and-error/learning by doing, it's exactly the opposite. Plan A buys us ten more years than we otherwise would've had to experiment on our AIs, understand how they work, and try out different regulatory approaches before we proceed to superintelligence.
Plan A does in fact bake in some pretty strong policy commitments. "Preventing progress is the trial-and-error friendly policy regime" would be an absurd position to take about policy stances toward technological progress in general. The point is not entirely without merit - AI development could outrun policy cycles (and arguably already is) - but the framing strikes me as pretending not to appreciate Séb's position, and that in turn make me less trusting toward the rest of the piece.
Thanks for this pushback. I think you are right about the top level that generally it was too angry, and decided to edit it to tone it down (but preserved the original and the diff here: https://docs.google.com/document/d/1i6Y9KTK2cyJ7y3XQtGY6f9F4nloRQr-rep8qSsicYUA/edit?tab=t.0 for transparency). I'm sorry about that.
Re your specific example, I still think that list is pretty misleading about what I actually think, so I do still want to set the record straight.
It does seem more valuable to litigate the economics disagreements on the object level, but we need clarity on the views themselves to do that productively. I think the econ supplement to ai 2040 is the best place we've done this so far: https://ai-2040.com/supplements/economics-of-plan-a
I'm mostly convinced by Plan A (given the techno-economic assumptions).
I still struggle with TRT as plausible without radical internationalisation. Financial incentives are too strong for companies, geoeconomic incentives too strong for states. The social contract would likely be unstable and therefore hard to enforce. AI companies already have strong levers with governments and are likely to use them to obstruct, dodge, cheat or FUD. We don't have enough time for it to become incontrovertible that cigarette smoke causes cancer.
I would thus trade algorithmic governance (inspired by Bioengeneering) with some open weight release (for Unis, independent researchers to participate in alignment research for positive long tail spillovers), ie (some) Research Transparency with (some) open weights.
You might be interested in other proposals we make such as filtered transparency (see here). I don't fully understand your comment but I do think that TRT is a bigger lift than filtered transparency. We recommended it because I think it is the best option, and I think it's viable if enough people push for it. But TRT is NOT a load bearing part of Plan A; more lax transparency proposals are very much consistent with it.
Hi Thomas, sorry this was not very clear,
I was trying to paint a rough sketch of "spillovers in context" most of which is done well in your scenario planning. My disagreement largely with the current rapidly developing "state-AI complex". The amount of private (and likely soon, public) capital, talent, state interest concentrated make it unlikely that TRT will be acceptable.
I would argue AI2040's linchpin is compute. Coordinate compute, open up almost everything else. Governments and companies want control of many other things, algorithms, talent, data etc. These will also be negotiable between companies and states. This means companies, states get to keep certain advantages in return for offers . Algorithms alongside compute are likely to be the most significant assets under negotiation & both will be part of asks and offers.
I will definitely have a look at filtered transparency, it may help anticipate some of the mechanics here.
In the case of AI, training code and training data is far more auditable than the resulting model weights, which are (for now) an uninterpretable pile of numbers.
I agree with this in the current regime, however as RL becomes an increasingly large share of training compute, I think it becomes increasingly important to access model weights.
Rollouts are a good starting point—though maybe these aren't even counted as "training data", since they can't be audited before training—but they don't really address concerns about how the RL generalized off-distribution, whether it selected for hidden reasoning/scheming, etc. In this setting, the actual policy and reward models would probably be more informative for auditing than just the training code and data.
The reason why China would go for Plan A is mostly that the status quo is terrible for them, and as the AIs get more capable and the US stays ahead, we expect them to wake up to this reality.
Seb’s assertion is that your consideration of the geopolitics of negotiations with China is naïve, and this response doesn’t seem to refute that. A negotiated settlement is only one of many options in China’s toolbox. Some potential considerations:
Getting China onboard for this proposal would be somewhere on the order of the greatest geopolitical coup in the history of the world. Treating it like an afterthought that will just naturally work itself out “because AGI/ASI” feels like skipping the hard part of the conversation. I would have wanted to see drastically less sci-fi speculation and drastically more engagement with on-the-ground facts of current geopolitical reality, since that’s what will actually make or break this plan.
Two days after publishing this post, we decided that the tone was unhelpfully combative. We stand by our views, but we made edits to make the tone less combative, and we removed two items from the list of false representations. You can see a diff here.
This criticism of AI 2040: Plan A by Séb Krier unfortunately seriously mischaracterizes our proposal. While we appreciate constructive criticisms of Plan A, such as the ones by Tom Davidson, Richard Ngo, and 1a3orn, we feel the need to correct the issues in Séb’s response. First, we’ll go over the specific false representations, and then we’ll give a point-by-point response.
False Representations
The exact opposite is true. Plan A is extremely iterative. In the status quo, there is trial and error, but ultimately companies aren’t going to choose the safer or more societally beneficial path, they are going to choose what the market wants. In Plan A there is much more time for AI companies to gain evidence and for governments to respond reasonably to the sweeping changes. Thanks to total transparency and broad deployment, all of this evidence is accessible to academics, independent researchers, and the public instead of being sealed away in the labs where only lab insiders can see it. Our plan maximizes learning and room to experiment.
Much of this is wrong. Specifically, we believe: (b) not all profits will accrue to labs,[1] (c) that Plan A would significantly decrease lab profits,[2] and (e) ‘de facto’ nationalization is not an optimal response.[3]
This is an out of context quotation to make us sound bad. You can see this quotation in context here. We note the absence of a global central planner as a constraint Plan A must work within. The Deal we propose does not create such a central planner, nor do we think this is a fundamental issue with the status quo.
See this quotation in context. We do not say, as is implied here, that these three possible outcomes are exhaustive. The whole point of writing Plan A was to increase the chance of a better outcome, namely a peaceful, multipolar world where the benefits of AI are shared widely.
Point-by-point response
We’ll now go through point-by-point and respond to everything.
We think this understates the originality of Plan A. Our policy recommendations are certainly not the same thing you've been hearing from the AI safety milieu since 2018. For instance, where else have you heard someone seriously propose Total Research Transparency?
Eisenhower said “Plans are worthless, but planning is everything” and we agree. Plan A as stated will obviously never happen exactly as we laid out. But the same was true of the plans for the Normandy invasion; reality quickly diverged from the plans, and the planners knew that this would happen. But trying to coordinate a massive effort like that without a plan would have been completely hopeless.
It is actually very difficult to design a coherent scenario that favors a particular outcome — over the course of writing Plan A, we often wrote up ideas that we thought were good, but then turned out to be inconsistent with some other aspect of the vision, thus forcing us to choose. Writing up scenarios like this is a huge amount of work, but we would nevertheless strongly recommend AGI companies such as GDM, OAI, and Anthropic invest the effort to actually try to articulate an internally coherent vision.
Re: blurring descriptive and normative, we do wish we could have done better here. We wrote "Plan A Assumptions" in part to clarify these issues, but it’s unfortunately an inherent difficulty of the format. In our view, scenario recommendations are the worst ways to make AI policy recommendations, except for all the other ways, which have much worse problems (such as allowing their authors to avoid making any real claims at all about the future, as is the norm in most AI policy papers).
As we say above, this is a mischaracterization of Plan A. We don’t propose a central planner; our proposal would spread out the power in several ways instead of concentrating it in a single AI, company, or government. There would be more companies at the frontier of AI development, more countries with power over how AI is governed, and more transparency so that the public can see what’s being done with AIs. Far from appointing a “cadre of elites” to control the fate of AI, Plan A enables a wider range of people to be involved in the conversation and have power over how it all goes down compared to the status quo.
As for our plan baking in too much and not leaving room for trial-and-error/learning by doing, it's exactly the opposite. Plan A buys us ten more years than we otherwise would've had to experiment on our AIs, understand how they work, and try out different regulatory approaches before we proceed to superintelligence. And thanks to total transparency and broad deployment, all of this evidence is accessible to academics, independent researchers, and the public instead of being sealed away in the labs where frontier AI employees can see it. Our plan maximizes learning and room to experiment.
Séb is underestimating how much power the government already has over AI. He objects to using network taps for verification when the government can already see what's going on in the datacenters if it wants to. Whenever he pleases, the President can hijack a company using the DPA, destroy its revenue source by blocking model deployment with export controls, and so forth. How exactly is Plan A giving them more power than they already have? The point of the network taps is to diffuse power by enforcing transparency — allowing the public and foreign governments to see what the US government can see in the status quo. Everyone gets to see the code that trains the AIs that are eating the whole economy, which is crucial for preventing extreme power concentration.
Point-by-point: (a) We've given detailed arguments for fast diffusion and economic transformation. (b) This is an exaggeration of our view, though we do think the AI companies will be the single biggest winners in an economy transformed by AI, capturing much more of the profits than, eg, workers or owners of capital other than semiconductors and robots. (c) This is not our view. Our policies — most notably including total research transparency — would dramatically cut AI company profits relative to the Plan D counterfactual.[4] (d) We've explained at length why we predict explosive growth soon, and Séb doesn't engage with our explanations at all. (e) We do not propose nationalizing the AI companies, as Séb would know if he had read our plan.
This misrepresents Plan A. "Solve the technical alignment problem, and then simply hand it to the UN to implement everywhere" is not at all what we call for. In our scenario, private AI companies iteratively solve the alignment problem while regulators set a low floor by banning the most obviously dangerous AI R&D practices. Plan A is the polycentric, competitive version of AI governance. Moreover, Plan A is the most plausible case where "muddying through" works out for humanity. In the other plans, you get a few months at most to try out different alignment strategies and make mistakes before you have to hand off to your AIs and pray they're in the basin of good deference. In Plan A, by contrast, you get to spend ten years muddling and figuring safety out on the fly.
The reason why China would go for Plan A is mostly that the status quo is terrible for them, and as the AIs get more capable and the US stays ahead, we expect them to wake up to this reality. Seb argues that the “heavy tax on AI persuasion” is absurd; this is an area where we actually don’t love our proposed solution. Indeed, one of my top recommendations for future work is for someone to try to operationalize this better.
Seb claims: “asymmetries and incentives will simply push the race into a more dangerous state-level one”. We disagree. The race dynamics post Plan A are much better and more manageable than what would have happened by default because the countries have the machinery to slow down in the face of danger and coordinate on safer paths through the tech tree.
What exactly are these questionable assumptions about the nature of AI? Apparently Séb thinks we can't speak of AIs having drives or being aligned/misaligned. But then how are we supposed to speak about AI behavior? It's not just renegade safety people who use mental language for AIs. AI practitioners within the frontier labs routinely talk about AI alignment, motivations, and so on. And they do so because it's useful and predictive to take the intentional stance with respect to AIs.
As for our supposedly unquestioned dogmas about unitary entities with drives of its own etc… If what you mean is that there can only be one AI with one set of goals and values, that's not what we predict. In our scenario there are many AIs with many different goals and values. But it’s unclear exactly how and why your view on AI goals differs from ours. We’ve written up our reasoning about AI goals and values here, though we believe it has been mostly obsoleted by now.
As for the “sudden uncontrollable discontinuous leap”...we don’t believe in this. Have you read AI 2027 or our takeoff model? Notice how continuous the curves are: we currently tend to expect a smooth intelligence explosion over the course of months or years, though we have heated disagreement within our team about the specifics.
First on whether Plan A is closer to mandating open source or to banning it, it helps to back up and remember that the key benefit of OS software is auditability. Instead of having to trust that some piece of software was written in the user's interest and is not backdoored, compromised, etc, the user can just directly audit the software and confirm it does what they want it to do. In the case of AI, training code and training data is far more auditable than the resulting model weights, which are (for now) an uninterpretable pile of numbers. So if you're a user who wants to confirm that your AI is working for you and not pursuing some covert agenda, total research transparency will do you a lot more good than mandatory open weights. Further, open weights are easier for a covert AI project or other bad actor to abuse than total research transparency, since it takes far less compute to train refusal out of a finished model than to train a helpful-only model from scratch using transparent source code. This is why we recommend TRT but not open weights.
It's true that TRT would neuter commercial incentives to develop new AI algorithms, but this is a feature, not a bug. A key principle of Plan A is limiting algorithmic progress because it's irreversible — once an algorithm is discovered, there's no wiping it from everyone's memories — and hard to keep from leaking to covert projects. RE: diffusion, we think that the natural diffusion of AIs that are as capable as the top human expert in every domain will be plenty fast enough. We think that trying to maximize capabilities and hence racing all the way to superintelligence for the sake of “diffusion” would be a terrible mistake. A key benefit of diffusion is that it will allow society to react and implement measures like those discussed in Plan A, but not if capabilities progress outpaces our reaction speed.
We see it as a pro rather than a con of Plan A that many of the requisite mechanisms are already being developed. We don't see why Séb presents this as an objection to the plan. As for the people who thought these things were never going to happen, we don’t know any such people.
About pausing at GPT-2, of course pausing at low levels of capability would reduce the rate of safety progress relative to pausing at higher capability levels. This is why we call for continued scaling to ~max controllable capabilities level before pausing. You can see our more detailed analysis of the pros and cons of scaling faster or slower here.
Conclusion
Overall, Séb’s piece is mostly inaccurate and fundamentally misunderstands what we are suggesting. It seems like in practice, Séb is advocating for something like Plan D, i.e., what happened in the race ending of AI 2027. We’ve presented many arguments for why we think this is terrible — it would pose unacceptably high AI takeover risk and massively concentrate power into AI companies or the US government.
There are some real problems Séb points out (e.g. our proposal to limit AI persuasion), but he doesn’t seem to acknowledge that these problems exist in his preferred world as well, except they are much worse because there is no slowdown and less transparency.
Still, for this reason and many others, Plan A is far from perfect. We wish that we had more specific and better proposals for many parts of Plan A, and welcome suggestions for improvements.
Nvidia is currently making large profits, and in Plan A we expect that to continue. We do think the AI companies will be the single biggest winners in an economy transformed by AI, capturing much more of the profits than, eg, workers or owners of capital other than semiconductors and robots.
Our policies—most notably (i) the massive AI capabilities slowdown and (ii) total research transparency—would dramatically cut AI company profits relative to the Plan D counterfactual in the short run. In the slightly longer run, profits are much greater than they would've been in Plan D because we would probably be dead.
The AI companies do not get nationalized in Plan A, and in fact the government’s ability to bully them goes down compared to today thanks to the transparency. (For example today the exec branch can decide to export control models, preventing them from being deployed externally, and the public isn’t able to see how the decision was made or whether the model in question really was more dangerous than competitor models for example.)
In the short run, that is. In the slightly longer run, unless we get lucky with alignment, company profits fall to zero in Plan D when everyone dies.