AI Safety Is Testing the Wrong Environment
The Lab Problem in AI Safety There's something that doesn't sit right with me about where alignment research happens. So much research, so many researchers, ideas, experiments, but the environment makes no sense. Almost all of it takes place in the same setting: one AI, one user, a chat interface,...
Jul 29