Notes on MoReBench: "Evaluating Procedural and Pluralistic Moral Reasoning in Language Models, More than Outcomes"
Notes pt. 4, decided on MoReBench!Paper Link: https://arxiv.org/abs/2510.16380Summary (What)AI systems are being increasingly used for decision making, but how are models actually making their decisions? MoReBench has 1000 moral scenarios, each with a rubric made by experts. MoREBench was created for both AI as an advisor and as an agent. The authors also created MoReBench-Theory: a selection of 150 scenarios under five major moral frameworks, aiming to test whether models could reason in accord...
Read full article →