AI Rights Aren't Safety-Neutral: A Quick Follow-Up to the Consciousness Cluster
TLDRThis short post is a quick write up of a short 3-day project I did as part of ARBOx. Taking inspiration from Chua's 'Consciousness Cluster' paper we decided to follow-up by asking what downstream behaviour changes we might observe if we fine-tuned/prompted a model to focus on legal rights and personhood (an increase in power-seeking and a decline in corrigibility). This post covers a short discussion of the results, limitations and methodology of what we did. In my personal opinion though, I...
Read full article →