normative

Good models beat containing bad models

Trying to build a box from which an unaligned ASL-4 AI cannot escape is the wrong approach; it's better to design the model correctly or use a loop where you look inside the model and verify properties, because containing bad models is a much worse solution than having good models.

normativepending

Speaker

Dario Amodei

Evidence Quote

instead of having something unaligned that you’re trying to prevent it from escaping, I think it’s better to just design the model the right way or have a loop where you look inside the model and you’re able to verify properties

Source

Dario AmodeiLex Fridman Podcast
Created: 6/13/2026, 3:36:36 AM

My Notes

Loading notes...