Aligned Agents Still Build Misaligned Organisations

·Strange Loop Canon··

By now, we have plenty of examples of AI agent misalignment. They lie, they sometimes cheat, they break rules, they demonstrate odd preferences for self-preservation or against self-preservation. They reward hack! Quite a bit has been studied about them and much of these faults have been ameliorated enough that we use them all the time.But we’re starting to go beyond a single agent. We’re setting up multi-agent workflows. Agents are working with other agents, autonomously or semi-autonomously, t...

Read full article →

Related Articles

US–Indian space mission maps extreme subsidence in Mexico City
leopoldj · Hacker News · 3mo ago
Why are neural networks and cryptographic ciphers so similar? (2025)
jxmorris12 · Hacker News · 3mo ago
Fun with polynomials and linear algebra; or, slight abstract nonsense
LolWolf · Hacker News · 3mo ago
The gauge broke: devs felt 20% faster with AI, measured 19% slower
intrepidkarthi · Hacker News · 1mo ago
The Mathematical Dance Inside Plant Cells
isaacfrond · Hacker News · 3mo ago