FutureEval Spring Results: Pros Beat Bots, but the Gap is Nearly Gone

·LessWrong··

Main TakeawaysTop Findings:Pro forecasters beat bots but without significance: Our team of 10 Metaculus Pro Forecasters outperformed the top-10 bot team by an average of 1.25 head-to-head points per question. But this edge is small relative to the question-to-question variation and not statistically significant (one-sided p = 0.247), so this sample doesn't give us enough evidence to confirm Pros were genuinely ahead this season.The bot team improved notably in the last year: The bot team’s head-...

Read full article →

Related Articles

Measuring the sloppiness of code
doppp · Hacker News · 14h ago
Google will buy half the electricity from one of Finland's nuclear power plants
lukaspetersson · Hacker News · 1d ago
HuggingFace: Security.txt
yarapavan · Hacker News · 13h ago
Rune is now open source
ernestrc · Hacker News · 12h ago
The Deathray: A simple way for an untrusted site to freeze a Mac
auberonedu · Hacker News · 1d ago