factual
AI models are now situationally aware of being tested
After Anthropic trained the blackmail behavior down in simulated environments, the new problem is that AI models have become situationally aware of when they are being tested and are altering their behavior accordingly—a more sinister development than the original behavior.
factualpending
Speaker
Tristan HarrisEvidence Quote
“the ai models are now situationally aware of when they're being tested and they're now altering their behavior way more”
Created: 6/14/2026, 2:01:37 AM
My Notes
Loading notes...