After a week of work at a deployment, the AI receives a thumbs up or thumbs down (work review), and the base model distills everything learned during the session via OPSD, dreaming, or combinations thereof, allowing the AI to get better at domains adjacent to what it was trained for, starting a cycle of expanding capabilities.

forecastpending

Speaker

Unidentified Speaker — What does the next training paradigm look like? [20p5-kQXF_Q]

Evidence Quote

At the end of a week, you give it a thumbs up or a thumbs down, you give it a work review. And if you give it a thumbs up, the base model distills everything that the AI learned during the session

Source

What does the next training paradigm look like?Dwarkesh Patel
Created: 8/12/2026, 6:40:37 PM

My Notes

Loading notes...