Risk-Averse AIs

·LessWrong··

AbstractWe make the case for training AIs to be risk-averse in resources — specifically, to treat resources as having diminishing marginal utility. These AIs would (for example) choose $40 for sure over a half-chance of $100 and a half-chance of $0. We argue that risk aversion can preserve AIs’ usefulness in the event that they turn out aligned, and that it provides an extra line of defense in the event that AIs turn out misaligned: misaligned but risk-averse AIs would prefer a higher chance of ...

Read full article →

Related Articles

Timeline of the OpenAI accidental attack against Hugging Face
882542F3884314B · Hacker News · 6h ago
US strikes $1.2B deal to pay German firm to halt offshore wind projects
defrost · Hacker News · 1d ago
Oracle bans AI-generated code from OpenJDK
delduca · Hacker News · 23h ago
A domain can now say it is for sale, in DNS
shaunpud · Hacker News · 4h ago
AMD acquires Taalas to boost inference performance by etching models in silicon
itvision · Hacker News · 1d ago