This is a special post for quick takes by Smaug123. Only they can create top-level comments. Comments here also appear on the Quick Takes page and All Posts page.
Anyone got any alignment-relevant problems they think might eventually be amenable to brute LLM usage, ideally ones which are probably not getting attention elsewhere? For example, my amateur understanding is that threat-resistant bargaining was a very interesting problem which Diffractor put a lot of work into, but which has stalled, not at any fundamental barrier but simply because our civilisation isn’t really trying to solve it. I’m happy to devote some convenient portion of a Max subscription of Claude and ChatGPT to get the frontier models working on that sort of question.