Single Forward Pass Evals on Fable, Opus 5, and GPT-5.6-Sol

·LessWrong··

This is a research update for an on-going replication of single-forward-pass evals done as part of the Second Look Fellowship. In following posts, we will run more comprehensive replications of previous work and release open source tooling for single forward pass eval elicitation. Code can be found here.tl;drWe replicate experiments from Greenblatt 2025 and Greenblatt 2026 on one baseline model from the original post, Opus 4.5. Our evaluations agree with the trends and quantitative values descri...

Read full article →

Related Articles

Hackers Got Inside a Flock Camera
driverdan · Hacker News · 13h ago
Apple Reference Image: A New Approach for Verified Photography
imwally · Hacker News · 1d ago
Training a 4B model to produce 81% faster query plans than Postgres
polyphilz · Hacker News · 8h ago
Xiaomi Mimo 2.6 live post-training dashboard
krackers · Hacker News · 6h ago
Nvidia announces native GPU programming in Rust
nonmaskable · Hacker News · 15h ago