A study on instability of LLM responses as a behavioral signature of self-Referential reports.

·LessWrong··

Introduction and Related workThe first person perspective of various experiences are subjective experiences. For Large language models, the study of subjective experiences was recently studied by Berg et al. (2025) who found out that self-referential prompting increases first person reports resembling subjective experience across GPT, Claude and Gemini. They also found out that reducing features associated with deception and roleplay increases the self-referential effect. Hahami et al. (2025) us...

Read full article →

Related Articles

F-Droid 2.0
daveoc64 · Hacker News · 15h ago
Two-tier encryption in the UK
ReturnoftheHack · Hacker News · 20h ago
Google’s Project Suncatcher to put ML infrastructure in space
xnx · Hacker News · 17h ago
Italian parliament votes for return to nuclear energy
geox · Hacker News · 1d ago
Toyota is taking the Corolla electric
cisc · Hacker News · 1d ago