What just happened? A retrospective of AI alignment

·LessWrong··

This sequence is about the last decade in AI alignment. Over five posts, it recounts the gradual transition from a field which treated alignment as a hard scientific problem, to a field which has largely abandoned the goal of deep, generalizable scientific progress in favor of iteratively improving existing systems and attempting to gain technological and political power. I also describe (in subsequent posts, which I'll upload over the next few weeks) how fear and (self-)deceptive reasoning made...

Read full article →

Related Articles

Italian parliament votes for return to nuclear energy
geox · Hacker News · 7h ago
Claude Opus 5.5
km144 · Hacker News · 1d ago
GPT-6 Astra has gained the ability to drive a car
plurby · Hacker News · 8h ago
UK military jamming other nations' satellites to defend itself, BBC told
thm · Hacker News · 6h ago
Claude Code reads AGENTS.md only when telemetry is on [fixed]
pszypowicz · Hacker News · 11h ago