Adversarial cryptography models from complexity theory can be applied to validate AI systems and test for unwanted properties like deception, representing an opportunity for academic computer science to contribute to AI safety.

factualpending

Speaker

David Nermberg

Evidence Quote

the adversarial models of cryptography can be applied to do validity testing for AI

Source

Sir Demis Hassabis on The Future of Knowledge | Institute for Advanced StudyInstitute for Advanced Study
Created: 8/11/2026, 7:33:14 AM

My Notes

Loading notes...