V&V takes on OpenAI’s long-horizon incidents by Yoav Hollander

·Nuno Sempere··

[Cross-posted from The Foretel­lix CTO Blog. Th­ese short takes try to put a ver­ifi­ca­tion-and-val­i­da­tion slant on AI-safety /​ al­ign­ment top­ics – they are not full treat­ments. I co-origi­nated cov­er­age-driven ver­ifi­ca­tion (CDV), and spent sev­eral decades do­ing ver­ifi­ca­tion of chips and AVs. See in­tro post for back­ground.]On July 20 and 21, OpenAI pub­lished two un­usu­ally can­did in­ci­dent re­ports: one about their in­ter­nal long-hori­zon model (the Erdős one) mis­be­hav...

Read full article →

Related Articles

Will Starship Ship 40 be recovered to shore?
Mqrius · Manifold Markets · 1d ago
When Is It Worth Personally Preparing for AI Disasters? by Ben_Norman
Ben_Norman · Nuno Sempere · 2d ago
On August 7th will Claude Opus 5 be above GPT 5.6 Sol on LLM agent arena?
Tetra · Manifold Markets · 3d ago
LLM-Assisted Proofreading A Study of the Effectiveness of 13 Free AI Detection Tools by Kariema El Touny
Kariema El Touny · Nuno Sempere · 4d ago
What can poor countries sell in the AI age? by Deena Mousa
Deena Mousa · Nuno Sempere · 5d ago