🧩 Philosophy 2d ago · Ben Wilson

FutureEval Spring Results: Pros Beat Bots, but the Gap is Nearly Gone

Less Wrong
View Channel →
FutureEval Spring Results: Pros Beat Bots, but the Gap is Nearly Gone
Source ↗ 👁 3 💬 0
Main TakeawaysTop Findings:Pro forecasters beat bots but without significance: Our team of 10 Metaculus Pro Forecasters outperformed the top-10 bot team by an average of 1.25 head-to-head points per question. But this edge is small relative to the question-to-question variation and not statistically significant (one-sided p = 0.247), so this sample doesn't give us enough evidence to confirm Pros were genuinely ahead this season.The bot team improved notably in the last year: The bot team’s head-

Comments (0)

Sign in to join the discussion

More Like This

Zizek at the Cricket Club
3:AM Magazine · 1d ago
How Different Are Philosophers? (guest post)
Daily Nous · 1d ago
Mini-Heap
Daily Nous · 1d ago
How to Change Your Life
The Marginalian · 1d ago
📰
Recommendations for People Getting into Technical AI Governance Research
LessWrong · 2d ago
📰
Personal statement on joining the OpenAI nonprofit board
LessWrong · 2d ago