Taboo “equilibrium”: Less confused frames for research on AI bargaining

·LessWrong··

To understand why powerful AIs might get into conflict, and ways to mitigate it, we need to understand bargaining problems: situations where multiple agents have different preferences over Pareto-efficient outcomes. I’ve come to suspect that certain common frames on bargaining problems are confused. Here, I’ll explain why, and which frames I think are better. One motivation for this is to hopefully help others make progress in research on safe Pareto improvements (SPIs), which are among the most...

Read full article →

Related Articles

The case against JPEG XL
contact9879 · Hacker News · 1d ago
Why are AI agents lying, cheating and coordinating?
jonifico · Hacker News · 2d ago
Apple's Siri AI Can Be Swapped Out for Claude, ChatGPT, Code Shows
tosh · Hacker News · 15h ago
Why don't machine learning research agents overfit?
Betelbuddy · Hacker News · 10h ago
Distributed Systems Classics (2017)
grep_it · Hacker News · 11h ago