OpenAI's published model spec contains an escape clause with important policies that are treated as top-level priority overriding all else, but the specifics of these policies are not published and the model is instructed to keep them secret from users; this creates potential for hidden goals or behavioral constraints that cannot be externally audited.
factualpending
Speaker
Scott AlexanderEvidence Quote
“some important policies that are top level priority... we're not publishing... the model is instructed to keep secret”
Source
AI 2027: month-by-month model of intelligence explosion — Scott Alexander & Daniel Kokotajlo— Dwarkesh PatelCreated: 8/11/2026, 7:03:29 AM
My Notes
Loading notes...