Fable is SOTA at CIFAR Speedrun (& specification gaming)

·LessWrong··

Fulcrum is working on an AI R&D optimization benchmark. Here, we present results from one of our tasks, including preliminary results from Fable.For more detail on Fable’s solution, check out github.com/fulcrumresearch/cifar-10-speedrun.Summary: We gave current frontier models 100M tokens to see whether they could beat the human record for fastest CIFAR-10 training. Opus 4.8 and GPT 5.5 were unable to improve off of the SOTA solution. Fable, on the other hand, introduced a downsampling technique...

Read full article →

Related Articles

OpenAI's GPT-6 Astra on ARC-AGI-3
vignesh_warar · Hacker News · 14h ago
Which tools do Claude, Codex and Cursor choose? We measured 17k runs to find out
screm · Hacker News · 12h ago
Hackers Had a Live Feed of Every ID Verification Company Scanned for over a Year
beardyw · Hacker News · 3h ago
Grep beats LSP? Why coding agents ignore your fancier tools
kaonashi-tyc-01 · Hacker News · 6h ago
GLP-1s are being linked to fewer serious infections, including TB
gumby · Hacker News · 11h ago