A draft honesty policy for credible communication with AI systems by Forethought

·Nuno Sempere··

This is a rough re­search note – we’re shar­ing it for feed­back and to spark dis­cus­sion. We’re less con­fi­dent in its meth­ods and con­clu­sions.ContextWe think that it would be very good if hu­man in­sti­tu­tions could cred­ibly com­mu­ni­cate with ad­vanced AI sys­tems. This could en­able pos­i­tive-sum trade be­tween hu­mans and AIs in­stead of con­flict that leaves ev­ery­one worse-off.[1] We want mod­els to be able to trust com­pa­nies when they make an hon­est offer or share in­for­ma­...

Read full article →

Related Articles

What is recursive self-improvement, and what would it mean to ban it? by sarahhw
sarahhw · Nuno Sempere · 22h ago
State of the Field: AI for Epistemics and Coordination by Ben_N
Ben_N · Nuno Sempere · 22h ago
Will a non-lab AI-only effort solve a Millennium Problem in 2026?
Bayesian · Manifold Markets · 1d ago
Will the U.S. use a nuclear weapon against Iran’s Pickaxe Mountain between November 4 and November 10, 2026?
Laurent27bis · Manifold Markets · 1d ago
How many people will attend EAG NYC 2026?
d · Manifold Markets · 1d ago