V&V takes on OpenAI’s long-horizon incidents

·LessWrong··

[Cross-posted from The Foretellix CTO Blog. These short takes try to put a verification-and-validation slant on AI-safety / alignment topics – they are not full treatments. I co-originated coverage-driven verification (CDV), and spent several decades doing verification of chips and AVs. See intro post for background.]On July 20 and 21, OpenAI published two unusually candid incident reports: one about their internal long-horizon model (the Erdős one) misbehaving during internal use, and one about...

Read full article →

Related Articles

Alphabet's cash burn raises alarm for Big Tech as AI spending climbs
1vuio0pswjnm7 · Hacker News · 7h ago
DARPA, U.S. Air Force fly AI-controlled F-16
r2sk5t · Hacker News · 7h ago
Everyone should know SIMD
WadeGrimridge · Hacker News · 1d ago
Cruller: Bun's Zig Runtime, Continued on Zig 0.16
Erenay09 · Hacker News · 15h ago
LG to ban residential proxies from smart TV apps
DemiGuru · Hacker News · 1d ago