How risky would it be to make powerful AI obey one or a few people?

·LessWrong··

It seems fairly likely that the first powerful AIs will be instruction-following rather than value-aligned, and will be controlled by a small number of people. So it makes sense to worry what individual people might do with such immense power. Here intuitions diverge and careful analysis is scarce. This post presents a debate between Seth Herd and cousin_it over how risky such a scenario would be. The debate ran under an unusual protocol. First we wrote our initial draft statements and sent them...

Read full article →

Related Articles

Dutch governments builds alternative for Microsoft based on NixOS
fjfaase · Hacker News · 16h ago
Revealing the details of how OpenAI agents hacked Hugging Face
specked-citrus · Hacker News · 3h ago
Google’s Project Suncatcher to put ML infrastructure in space
xnx · Hacker News · 1d ago
Platform-independent SIMD in Go
yurivish · Hacker News · 12h ago
F-Droid 2.0
daveoc64 · Hacker News · 1d ago