V&V takes on OpenAI’s long-horizon incidents by Yoav Hollander

·Nuno Sempere··

[Cross-posted from The Foretel­lix CTO Blog. Th­ese short takes try to put a ver­ifi­ca­tion-and-val­i­da­tion slant on AI-safety /​ al­ign­ment top­ics – they are not full treat­ments. I co-origi­nated cov­er­age-driven ver­ifi­ca­tion (CDV), and spent sev­eral decades do­ing ver­ifi­ca­tion of chips and AVs. See in­tro post for back­ground.]On July 20 and 21, OpenAI pub­lished two un­usu­ally can­did in­ci­dent re­ports: one about their in­ter­nal long-hori­zon model (the Erdős one) mis­be­hav...

Read full article →

Related Articles

New Research: Climate Mitigation is Overlooked by EA by Dan Stein
Dan Stein · Nuno Sempere · 1d ago
Will any non-astronaut be "commuting to the moon" before 2040?
Panfilo · Manifold Markets · 1d ago
Lexical Filtering: Deciding under Indeterminacy with Lexically Ordered Values by Kuutti Lappalainen
Kuutti Lappalainen · Nuno Sempere · 2d ago
Research Report: A genetic basis for potential nociception in lac bugs (Kerria lacca; Hemiptera: Kerriidae) by Meghan Barrett
Meghan Barrett · Nuno Sempere · 2d ago
Will at least one U.S. state enact a binding statewide moratorium on new hyperscale data centers before 2027?
Kakonomics · Manifold Markets · 3d ago