How risky would it be to make powerful AI obey one or a few people?

·LessWrong··

It seems fairly likely that the first powerful AIs will be instruction-following rather than value-aligned, and will be controlled by a small number of people. So it makes sense to worry what individual people might do with such immense power. Here intuitions diverge and careful analysis is scarce. This post presents a debate between Seth Herd and cousin_it over how risky such a scenario would be. The debate ran under an unusual protocol. First we wrote our initial draft statements and sent them...

Read full article →

Related Articles

England set to be one of the first countries to eliminate hepatitis C
stevekemp · Hacker News · 7h ago
Woman Pulled over at Gunpoint Twice After Flock Camera Glitch
cdrnsf · Hacker News · 2h ago
Stealing Reasoning Traces from Proprietary LLM APIs
quantumgarbage · Hacker News · 6h ago
Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
riordan · Hacker News · 1d ago
London Underground begins scanning passengers' faces
BlueBerry2001 · Hacker News · 10h ago