When does an LLM’s model of you affect its behaviour?

·LessWrong··

Disclaimer: figures in this post are edited by ChatGPTIn my last post, I looked at what makes LLMs form opinions of their users: gender, age, socioeconomic status, education, and mood. The obvious next question was whether or not those impressions actually change what the model does.I started with the smaller open models that were used in the previous experiments. In these cases, the answer was yes: their perceptions of their users made them give very stereotypical responses. For example, steeri...

Read full article →

Related Articles

Hackers Got Inside a Flock Camera
driverdan · Hacker News · 14h ago
Apple Reference Image: A New Approach for Verified Photography
imwally · Hacker News · 1d ago
Training a 4B model to produce 81% faster query plans than Postgres
polyphilz · Hacker News · 8h ago
Xiaomi Mimo 2.6 live post-training dashboard
krackers · Hacker News · 7h ago
Nvidia announces native GPU programming in Rust
nonmaskable · Hacker News · 16h ago