Safety Cases We Can Check Together
Recently I’ve been invited by SpaceXAI to help them release the “For You” recommendation engine on GitHub, where simply releasing the running weights alone could lead to fraud-related risks, and would not even pin which system was released.Since we set up Taiwan’s AI Evaluation Center in 2023, I’ve been thinking about how to bring structured transparency to AI testing and public replay of those tests, in order to achieve all four CROPS values (capture-resistant, open, private, and secure) withou...
Read full article →