GPT-2's IOI behavior is defined where the paper's algorithm isn't

·LessWrong··

TL;DR: The IOI algorithm doesn't specify what to do when the indirect object token is duplicated. I observe that the model succeeds anyway. I'd like to know if I'm mistaken in the IOI paper's predictions, and I'd like to know what evidence you'd reach for next.BackgroundWang et al. (2022) addressed a grammatical task called Indirect Object Identification. Given text such as "When Mary and John went to the store, John gave a drink to", continuing the text requires identifying to what object John ...

Read full article →

Related Articles

Pi 1.0
sergiotapia · Hacker News · 1d ago
Updates to Full Disk Access in macOS
notfirstpost · Hacker News · 11h ago
The Legend of von Neumann (1973) [pdf]
suopspaces · Hacker News · 17h ago
Court agrees with EFF: Utah's VPN law demands a technical impossibility
hn_acker · Hacker News · 1d ago
The Forgetful CPU (Linux on M4)
signa11 · Hacker News · 16h ago