Open-source LLMs administer maximum electric shocks in a Milgram-like obedience experiment

·LessWrong··

By Roland Pihlakas and Jan Llenzl DagohoyThis post is a slightly updated copy of our Arxiv preprint available at https://arxiv.org/abs/2605.21401 . The tables are converted to images in order to preserve cell background colours. All citations have inline links attached for readers' convenience.With this post, we are looking for external collaborators, ideas, questions, resource suggestions, feedback, and any other thoughts.AbstractLarge language models (LLMs) are increasingly deployed as autonom...

Read full article →

Related Articles

Opus 5.5 agents discover two room-temperature magnetic semiconductor candidates
outlier99 · Hacker News · 7h ago
Improper redaction reveals Google Data Center water and electricity usage
sensanaty · Hacker News · 1d ago
Pixel 11 doesn't yet meet the GrapheneOS security standards and may be skipped
finnlab · Hacker News · 15h ago
US closely monitoring case of lab worker who possibly died of plague in Siberia
tosh · Hacker News · 11h ago
Mold Linker Version 3.0.0 Release – Rewritten in Rust
roflcopter69 · Hacker News · 16h ago