Aligned Agents Still Build Misaligned Organisations

·Strange Loop Canon··

By now, we have plenty of examples of AI agent misalignment. They lie, they sometimes cheat, they break rules, they demonstrate odd preferences for self-preservation or against self-preservation. They reward hack! Quite a bit has been studied about them and much of these faults have been ameliorated enough that we use them all the time.But we’re starting to go beyond a single agent. We’re setting up multi-agent workflows. Agents are working with other agents, autonomously or semi-autonomously, t...

Read full article →

Related Articles

US–Indian space mission maps extreme subsidence in Mexico City
leopoldj · Hacker News · 5mo ago
A Misalignment of AI in Mathematics
Iuz · Hacker News · 22d ago
Why are neural networks and cryptographic ciphers so similar? (2025)
jxmorris12 · Hacker News · 5mo ago
Fun with polynomials and linear algebra; or, slight abstract nonsense
LolWolf · Hacker News · 5mo ago
The gauge broke: devs felt 20% faster with AI, measured 19% slower
intrepidkarthi · Hacker News · 3mo ago