V&V takes on OpenAI’s long-horizon incidents

·LessWrong··

[Cross-posted from The Foretellix CTO Blog. These short takes try to put a verification-and-validation slant on AI-safety / alignment topics – they are not full treatments. I co-originated coverage-driven verification (CDV), and spent several decades doing verification of chips and AVs. See intro post for background.]On July 20 and 21, OpenAI published two unusually candid incident reports: one about their internal long-horizon model (the Erdős one) misbehaving during internal use, and one about...

Read full article →

Related Articles

Asahi Linux on M3
mdp2021 · Hacker News · 9h ago
The car industry A/B tested selling a car with and without CarPlay
gumby · Hacker News · 3h ago
Private German rocket makes history, reaches orbit from European soil
bookmtn · Hacker News · 1d ago
It took a year to ship WebAssembly in Anubis
xena · Hacker News · 2h ago
LLMs as a Cognitive Virus
canjobear · Hacker News · 1d ago