A preliminary experiment regarding consistency as a measure of conceptual abilities in language models

·LessWrong··

Cross-posting from my coworker Caspar Oesterheld's blog which I think is great and generally not well known.I’ve recently been working a little on whether consistency across different questions can be used as a measure of (and perhaps ultimately as a training target for) philosophical competence. I’m in the process of writing up the results into a paper. I’m here reporting results from a small, preliminary experiment that I ran late last year. I’ll leave a more careful discussion (with a proper ...

Read full article →

Related Articles

We got admin access to Baseten's production GitHub in 25 minutes
bearsyankees · Hacker News · 9h ago
Building a Linux GPU Driver for the M4 Mac Mini in One Month
ADevWithAnIdea · Hacker News · 8h ago
Show HN: An e-ink frame that hears birds and draws them as 1800s illustrations
arnemunthekaas · Hacker News · 15h ago
America's Driver's License Breach Is a National Security Disaster
hn_acker · Hacker News · 12h ago
How much oil-market buffer is left?
mcone · Hacker News · 8h ago