Study 2 Registration: Exploring representational counterparts of welfare-relevant indicators under post-training quantization

·LessWrong··

Epistemic status: this is intended as a higher-quality preregistration for the next study in a series originally described in "Does post-training quantization change welfare-relevant indicators in open-weight language models?"; I have tried to avoid making any claims I cannot support from the results of Study 1 and to be entirely clear about which statements here are speculative, but if any statement is ambiguous you should assume it is hypothesis rather than conclusion.We ask whether welfare-re...

Read full article →

Related Articles

Hackers Got Inside a Flock Camera
driverdan · Hacker News · 14h ago
Apple Reference Image: A New Approach for Verified Photography
imwally · Hacker News · 1d ago
Training a 4B model to produce 81% faster query plans than Postgres
polyphilz · Hacker News · 9h ago
Nvidia announces native GPU programming in Rust
nonmaskable · Hacker News · 16h ago
Xiaomi Mimo 2.6 live post-training dashboard
krackers · Hacker News · 7h ago