LeCun's proposed JEPA (Joint Embedding Predictive Architecture) predicts representations of video rather than pixels, learning in abstract terms what may happen as a consequence of actions; this approach combined with world models and cost functions could enable planning and safe, controllable AI systems.

normativepending

Speaker

Yan LeCun

Evidence Quote

basically instead of predicting the pixels in the in the video we predict a representation of the pixels in that video

Source

AI: Grappling with a New Kind of Intelligence | World Science FestivalWorld Science Festival
Created: 8/10/2026, 10:49:40 PM

My Notes

Loading notes...