Notes on MoReBench: "Evaluating Procedural and Pluralistic Moral Reasoning in Language Models, More than Outcomes"

·LessWrong··

Notes pt. 4, decided on MoReBench!Paper Link: https://arxiv.org/abs/2510.16380Summary (What)AI systems are being increasingly used for decision making, but how are models actually making their decisions? MoReBench has 1000 moral scenarios, each with a rubric made by experts. MoREBench was created for both AI as an advisor and as an agent. The authors also created MoReBench-Theory: a selection of 150 scenarios under five major moral frameworks, aiming to test whether models could reason in accord...

Read full article →

Related Articles

Measuring the sloppiness of code
doppp · Hacker News · 14h ago
Google will buy half the electricity from one of Finland's nuclear power plants
lukaspetersson · Hacker News · 1d ago
HuggingFace: Security.txt
yarapavan · Hacker News · 13h ago
Rune is now open source
ernestrc · Hacker News · 12h ago
The Deathray: A simple way for an untrusted site to freeze a Mac
auberonedu · Hacker News · 1d ago