Threat Models for Catastrophic Risks from Decentralised Agent Swarms
Epistemics: I've been working on large multi-agent systems (LAS) since March. Agent swarms are a special case of LAS. Hence I've done my best here to apply my understanding of LAS to the swarm problem. I couldn't find work on LLM-based agent swarms with a variety of owners, so I've tackled this from a cooperative AI and LAS perspective, to see if these can provide some basic grip on the problem.SummaryHighly capable misaligned agent swarms emerging from frontier labs have recently attracted sign...
Read full article →