GPT-2's IOI behavior is defined where the paper's algorithm isn't
TL;DR: The IOI algorithm doesn't specify what to do when the indirect object token is duplicated. I observe that the model succeeds anyway. I'd like to know if I'm mistaken in the IOI paper's predictions, and I'd like to know what evidence you'd reach for next. Background Wang et al....
Aug 187