AI Alignment at Which Abstraction Level?

·LessWrong··

This post is crossposted from my Substack, Structure and Guarantees, where I explore how formal verification and related ideas might scale to more complex intelligent systems. Here I start from the observation that AI alignment usually takes for granted a privileged role for human values, without asking why values should come from that particular abstraction level. I then give an independent engineering reason to worry about putting humans at the center of an alignment specification, drawing on ...

Read full article →

Related Articles

Xiaomi: New CPU matches Apple cores single threaded, much faster multithreaded
tosh · Hacker News · 1d ago
MS Paint and Photos inivisibly watermark even locally generated output with GUID
ComputerGuru · Hacker News · 1d ago
IPFS Maintainers Winding Down
iand · Hacker News · 23h ago
LLMs could control their host machines by exploiting inference engines
zdw · Hacker News · 20h ago
HelloAssembly The smallest possible complete Windows application
Bluestein · Hacker News · 3h ago