causal
Models reason well when allowed to think out loud
Language models are not fundamentally bad at multi-step reasoning; they are bad at mental multi-step reasoning when not allowed to think out loud, but when allowed to think out loud they are quite good, and this will improve significantly with better models and special training.
causalpending
Speaker
Ilya SutskeverEvidence Quote
“they are bad at mental multistep reasoning when they are not allowed to think out loud. But when they are allowed to think out loud, they're quite good.”
Source
Ilya Sutskever (OpenAI Chief Scientist) — Why next-token prediction could surpass human intelligence— Dwarkesh Patel PodcastCreated: 6/13/2026, 3:26:51 AM
My Notes
Loading notes...