For Claude, capability and dispreferring CDT are the ~same thing. Much more so than for GPT.

·LessWrong··

We've previously reported that decision-theoretic capabilities and favoring EDT/generalised-one-boxing over CDT correlate in LLMs (both measured by DTBench). (Note that EDT, for the most part, doesn't come apart from FDT / UDT on DTBench.[1]) Anthropic also replicate the same finding in their Opus 4.7 and Fable 5 model cards.We recently noticed something funny: Capabilities and preference against CDT answers basically perfectly for Anthropic models. This holds whether you measure capabilities us...

Read full article →

Related Articles

Qwen3.8 27B scores 52 on Artificial Analysis
anana_ · Hacker News · 10h ago
India has paved the way for charging merchants a fee on UPI transactions
monkey_monkey · Hacker News · 8h ago
AI-Generated GitHub Copilot “Autofix” Allowed Compromise of Snowflake's Jira
galnagli · Hacker News · 13h ago
A Preview of DuckDB v2.0
ibotty · Hacker News · 13h ago
Self hosted email continues to steeply decline
minusf · Hacker News · 16h ago