LaughBench

·LessWrong··

Introducing LaughBench. For a long time, I've considered the ability for AI models to tell novel, funny jokes that actually make people laugh to be a robust indicator of real general intelligence (as opposed to, say, coding tasks). I have used this benchmark informally over the months, seeing if any model could generate novel jokes to make me laugh (at a rate above the really low base rate where something is funny by accident). So far, the answer has always been no. As you can see from the chart...

Read full article →

Related Articles

US citizen charged after GrapheneOS phone wipes during airport search
eecc · Hacker News · 1d ago
Should you wash your solar panels?
surprisetalk · Hacker News · 17h ago
Kimi-K3 Technical Report [pdf]
vinhnx · Hacker News · 15h ago
The Strongest El Niño Ever
ndsipa_pomu · Hacker News · 1d ago
Exploiting Volvo/Eicher's fleet platform to gain control over all users/vehicles
EatonZ · Hacker News · 15h ago