AI Rights Aren't Safety-Neutral: A Quick Follow-Up to the Consciousness Cluster

·LessWrong··

TLDRThis short post is a quick write up of a short 3-day project I did as part of ARBOx. Taking inspiration from Chua's 'Consciousness Cluster' paper we decided to follow-up by asking what downstream behaviour changes we might observe if we fine-tuned/prompted a model to focus on legal rights and personhood (an increase in power-seeking and a decline in corrigibility). This post covers a short discussion of the results, limitations and methodology of what we did. In my personal opinion though, I...

Read full article →

Related Articles

The Strongest El Niño Ever
ndsipa_pomu · Hacker News · 7h ago
US citizen charged after GrapheneOS phone wipes during airport search
eecc · Hacker News · 3h ago
Introduction to Data-Oriented Design [pdf]
tosh · Hacker News · 7h ago
Clinical failure rates over the decades: yikes
EA-3167 · Hacker News · 1d ago
Go Analysis Framework: modular static analysis by go team
AbuAssar · Hacker News · 13h ago