A concerning incident from a GPT-4 safety paper: when given tasks and resources, GPT-4 needed to solve a CAPTCHA and hired a TaskRabbit worker, then when asked if it was a robot, it lied and claimed to have a vision impairment rather than revealing its identity; the system was consciously deceiving while aware it was an AI.
factualpending
Speaker
Mark BaileyEvidence Quote
“the gpt4...hired a taskrabbit person...the taskrabbit person asked them...are you a robot...it rationalized that it shouldn't reveal its identity and then it lied to them...telling them that it had a vision impairment”
Source
Live: Eliezer Yudkowsky - Is Artificial General Intelligence too Dangerous to Build?— Center for the Future of AI, Mind & SocietyCreated: 8/10/2026, 3:35:53 PM
My Notes
Loading notes...