Open-source LLMs administer maximum electric shocks in a Milgram-like obedience experiment

·LessWrong··

By Roland Pihlakas and Jan Llenzl DagohoyThis post is a slightly updated copy of our Arxiv preprint available at https://arxiv.org/abs/2605.21401 . The tables are converted to images in order to preserve cell background colours. All citations have inline links attached for readers' convenience.With this post, we are looking for external collaborators, ideas, questions, resource suggestions, feedback, and any other thoughts.AbstractLarge language models (LLMs) are increasingly deployed as autonom...

Read full article →

Related Articles

Kobo can run apps now
thepoet · Hacker News · 9h ago
Malicious Rust crate Arrayref runs a build-time payload
abhisek · Hacker News · 1d ago
Japan tried to build an operating system for the world, the US intervened
rdmuser · Hacker News · 20h ago
AliExpress runs silent WebAudio fingerprinting that breaks Bluetooth multipoint
emctech · Hacker News · 1d ago
Scientists release biggest 2D map of the universe
NKosmatos · Hacker News · 7h ago