x
Towards deployment-time misalignment continuation evals: lessons from recent loss of control incidents — LessWrong