Patterns and problems in emerging multiagent systems (Anthropic, Frontier Red Team)

·LessWrong··

Linkpost for some new Anthropic research on how agents coordinate (or don't). Not too long, pretty interesting. For example: The jist of the report is that Mythos 5 does way better at coordination than previous models across a few scenarios. For example, when multiple Mythos are given conflicting goals for a single shared codebase, they eventually realize the other agents aren't hostile:(...) we observe an emergent behavior where the agents propose and run a tournament for application performanc...

Read full article →

Related Articles

DeepSeek V4 Pro 0813
explosion-s · Hacker News · 15h ago
Someone is running mass vulnerability scans, spoofing AI bots like ClaudeBot
gavinhking · Hacker News · 17h ago
We tracked down the 16-year-old WAL-reset SQLite bug
ropbear · Hacker News · 17h ago
London Underground begins scanning passengers' faces
BlueBerry2001 · Hacker News · 1d ago
England set to be one of the first countries to eliminate hepatitis C
stevekemp · Hacker News · 1d ago