LLMs perform autoregressive token prediction by generating probability distributions over a finite token dictionary and sampling from them sequentially, which is fundamentally different from the optimization-based inference required for planning with world models.
definitionpending
Speaker
Yann LeCunEvidence Quote
“So you have a sequence of discrete tokens, and you train a system to predict the next token in a sequence from enormous amounts of data.”
Created: 8/12/2026, 10:03:11 PM
My Notes
Loading notes...