I'm very impressed by the proposed Total Research Transparency. I actually found it appealing even beyond the many reasons mentioned in the plan. It takes advantage of key properties of the current training paradigm, and this is actually desirable, because these properties are likely to remain in future AI systems:
What is the correct calculation currently? There's a good number of new people and orgs now working on alignment seriously, I think more spheres of research intellectually outside MIRI and EA are important (like Institute for Responsible Superintelligence). Should capabilities keep going to get even more attention and talent? Possible to argue that the vast majority of the world (including elites and technical people that in theory will help address the problem) still doesn't take the problem seriously.
Ilya Sutskever said on the Dwarkesh pod - "I place mor... (read more)