Differential acceleration of alignment-relevant capabilities is a bad bet

·LessWrong··

There is an idea floating around in the rough shape of "we need to accelerate capabilities that are differentially useful for safety research so AIs can help us make the future go better." The capabilities targeted are typically things bottlenecking alignment research, such as philosophical or conceptual reasoning.I feel nervous about this for two reasons. The first is that it's plausible that AI safety and AI R&D are bottlenecked by many of the same factors: AIs have poor epistemics, are bad at...

Read full article →

Related Articles

Formalizing Fermat's Last Theorem
jlebar · Hacker News · 2h ago
Hackers Had a Live Feed of Every ID Verification Company Scanned for over a Year
beardyw · Hacker News · 13h ago
US Military disables ad trackers on troops' phones
tencentshill · Hacker News · 7h ago
Solving the Jane Street reverse engineering challenge
anitil · Hacker News · 10h ago
Google AI Mode shows same products 21.6% more expensive than traditional search
DeepLogin · Hacker News · 8h ago