A study on instability of LLM responses as a behavioral signature of self-Referential reports.

·LessWrong··

Introduction and Related workThe first person perspective of various experiences are subjective experiences. For Large language models, the study of subjective experiences was recently studied by Berg et al. (2025) who found out that self-referential prompting increases first person reports resembling subjective experience across GPT, Claude and Gemini. They also found out that reducing features associated with deception and roleplay increases the self-referential effect. Hahami et al. (2025) us...

Read full article →

Related Articles

Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
riordan · Hacker News · 18h ago
Mistral Patent for “Code implemented tool calls”
theanonymousone · Hacker News · 14h ago
Kinney Drugs pulls back AI phone assistant after hundreds of customer complaints
kotaKat · Hacker News · 13h ago
Study links GLP-1 drugs to bigger jump in women's employment than a degree
metadat · Hacker News · 12h ago
Show HN: Needle2: 14MB agentic LLM for phones, wearables, smart home and robots
HenryNdubuaku · Hacker News · 10h ago