LLMs perform autoregressive token prediction by generating probability distributions over a finite token dictionary and sampling from them sequentially, which is fundamentally different from the optimization-based inference required for planning with world models.

definitionpending

Speaker

Yann LeCun

Evidence Quote

So you have a sequence of discrete tokens, and you train a system to predict the next token in a sequence from enormous amounts of data.

Source

Yann LeCun: Special Lecture on AI and World ModelsAl-Khwarizmi Applied Mathematics Webinar
Created: 8/12/2026, 10:03:11 PM

My Notes

Loading notes...