x
Notes on MoReBench: "Evaluating Procedural and Pluralistic Moral Reasoning in Language Models, More than Outcomes" — LessWrong