Risk-Averse AIs

·LessWrong··

AbstractWe make the case for training AIs to be risk-averse in resources — specifically, to treat resources as having diminishing marginal utility. These AIs would (for example) choose $40 for sure over a half-chance of $100 and a half-chance of $0. We argue that risk aversion can preserve AIs’ usefulness in the event that they turn out aligned, and that it provides an extra line of defense in the event that AIs turn out misaligned: misaligned but risk-averse AIs would prefer a higher chance of ...

Read full article →

Related Articles

Claude Opus 5.5
km144 · Hacker News · 3h ago
I asked Meta’s Muse for its filesystem and it sent me 6.8GB
Aeroi · Hacker News · 4h ago
AMD's random number generator can't generate a 0?
BruceEel · Hacker News · 11h ago
There's a high chance of devices being sold with GrapheneOS preinstalled in 2027
Cider9986 · Hacker News · 3h ago
NASA’s Mars Sample Return mission is dead
Muhammad523 · Hacker News · 1d ago