Your Brain Has an Attack Surface

·LessWrong··

About a year ago, I began transitioning from software engineering to AI safety research. I was drawn into this by a question that arose while building runtime security for software systems: how do you impose constraints on a system you can’t fully observe? In AI safety, this question is at the very core: if we can’t reliably control how AI systems communicate and coordinate with each other, we can’t impose any other security properties on them. Since then, I’ve completed the AGI Strategy course ...

Read full article →

Related Articles

GLM-5.3 is now open-weight
jeudesprits · Hacker News · 10h ago
EPA says power for data centers can sidestep pollution laws
Levitating · Hacker News · 12h ago
Just the rumour of a bug is enough to find an exploit these days
avsm · Hacker News · 9h ago
Pentagon's blacklisting of Anthropic was unlawful, US judge rules
softwaredoug · Hacker News · 14h ago
Saving 100 terabytes of memory by optimizing 1.1.1.1's DNS cache
TangerineDream · Hacker News · 1d ago