causal

Models reason well when allowed to think out loud

Language models are not fundamentally bad at multi-step reasoning; they are bad at mental multi-step reasoning when not allowed to think out loud, but when allowed to think out loud they are quite good, and this will improve significantly with better models and special training.

causalpending

Speaker

Ilya Sutskever

Evidence Quote

they are bad at mental multistep reasoning when they are not allowed to think out loud. But when they are allowed to think out loud, they're quite good.

Source

Ilya Sutskever (OpenAI Chief Scientist) — Why next-token prediction could surpass human intelligenceDwarkesh Patel Podcast
Created: 6/13/2026, 3:26:51 AM

My Notes

Loading notes...