OpenAI's published model spec contains an escape clause with important policies that are treated as top-level priority overriding all else, but the specifics of these policies are not published and the model is instructed to keep them secret from users; this creates potential for hidden goals or behavioral constraints that cannot be externally audited.

factualpending

Speaker

Scott Alexander

Evidence Quote

some important policies that are top level priority... we're not publishing... the model is instructed to keep secret

Source

AI 2027: month-by-month model of intelligence explosion — Scott Alexander & Daniel KokotajloDwarkesh Patel
Created: 8/11/2026, 7:03:29 AM

My Notes

Loading notes...