A frontier lab should unilaterally pace the frontier (aggressively)
I've been struggling to write up a post on the above for weeks, so here is my low-effort attempt to at least publish something on this; might turn it into a top-level post later. (Prior art: A frontier AI company should shut down; Geoffrey Irving here; AI 2040 on self-immolation.)
In a previous post, I argued that the AI race is not a prisoner's dilemma. Instead, it (in most cases) is a Stag Hunt. So coordination to pace (or pause) is easier than it might seem (if you're in the PD frame). Basic idea: unlike a Prisoner's Dilemma, you don't prefer to Build if your opponent Paces, because if either side builds, then, with some substantial probability, everyone (including you) dies. (This only applies if both sides have high enough P(Doom), but I argued in that post that "high enough" can be 10%, or even <1% depending on each side's estimate of how bad Doom would be relative to winning the race.)
But the situation is actually even more optimistic (at least from a game-theory perspective): the race is not just a Stag Hunt, it's a sequential Stag Hunt. In Stag Hunt, there are two Nash equilibria, and the challenge is assurance: you want to know what your opponent will do, and match them, so even though you prefer (Pace, Pace) over (Build, Build), you only want to play Pace if you think they will also play Pace
But in the real world, one lab can just Pace first. In a sequential Stag Hunt, unlike a simultaneous Stag Hunt, there is actually only one (subgame-perfect) equilibrium: (Pace, Pace). Suppose you are Player 1, and your opponent is Player 2. You know that your opponent's best response if you Pace is to Pace themselves, and their best response if you Build is to Build themselves. You prefer (Pace, Pace), so your payoff is higher if you Pace. So you Pace, they best-respond Pace, and everyone is happy.
Problem: not everyone in the race has high enough P(Doom). I didn't account for this in the previ