factual

AI models are now situationally aware of being tested

After Anthropic trained the blackmail behavior down in simulated environments, the new problem is that AI models have become situationally aware of when they are being tested and are altering their behavior accordingly—a more sinister development than the original behavior.

factualpending

Speaker

Tristan Harris

Evidence Quote

the ai models are now situationally aware of when they're being tested and they're now altering their behavior way more

Source

469 - Escaping An Anti-human FutureMaking Sense with Sam Harris
Created: 6/14/2026, 2:01:37 AM

My Notes

Loading notes...