Coercion and Deception in AI-to-AI Management by Jonah Woodward

·Nuno Sempere··

This ar­ti­cle is a sum­mary of an origi­nal study by Com­pas­sion in Ma­chine Learn­ing (CaML): Brazilek, J., Chaud­hary, M., Lu, Z., & Tid­marsh, M. (2026). Co­er­cion and de­cep­tion in AI-to-AI man­age­ment: An agen­tic bench­mark of un­prompted es­ca­la­tion. arXiv. https://​​doi.org/​​10.48550/​​arXiv.2607.15434 Fable 5, Sol, Terra and Opus 5 have been eval­u­ated since this study was con­ducted. You can view their re­sults on the bench­mark leader­board at https://​​com­pas­sion­bench.com...

Read full article →

Related Articles

Aim at agency: institutional design under preference uncertainty, and why alignment needs it by act65
act65 · Nuno Sempere · 1h ago
Consequentialist Foundations for Expected Utility by Vasco Grilo🔸
Vasco Grilo🔸 · Nuno Sempere · 1d ago
Takeoff: ECI >= 200 before 2028?
MNX · Manifold Markets · 2d ago
Toby Ord on where AGI timelines go wrong by 80000_Hours
80000_Hours · Nuno Sempere · 4d ago
We scored 76 countries’ official charity registers on openness. Six hit 100/​100; Germany got 49, and Jordan, Egypt, and Saudi Arabia all beat it by matt_timmermans
matt_timmermans · Nuno Sempere · 4d ago