How's it going? Reinforcement learning in language models recruits a functional welfare axis
In collaboration with David Chalmers and Pavel Izmailov. Work done at NYU. Andy wrote this summary of the paper, which you can find in full on the website, or, if you insist on a PDF, arXiv. Introduction We know that language models work in a vast and shadowy landscape of...
May 3029