Adversarial cryptography models from complexity theory can be applied to validate AI systems and test for unwanted properties like deception, representing an opportunity for academic computer science to contribute to AI safety.
factualpending
Speaker
David NermbergEvidence Quote
“the adversarial models of cryptography can be applied to do validity testing for AI”
Source
Sir Demis Hassabis on The Future of Knowledge | Institute for Advanced Study— Institute for Advanced StudyCreated: 8/11/2026, 7:33:14 AM
My Notes
Loading notes...