Another more speculative idea for addressing sample efficiency is 'dreaming'—where AIs build a good simulation of reality to rehearse new skills, try alternative strategies, and reinforce what works, allowing AIs to experience orders of magnitude more simulated samples in the same wall-clock time.

forecastpending

Speaker

Unidentified Speaker — What does the next training paradigm look like? [20p5-kQXF_Q]

Evidence Quote

If the AI can build a good simulation of reality against which to rehearse new skills, or try alternative strategies and reinforce what actually works, then AIs could experience orders of magnitude more simulated samples in the same wall-clock time.

Source

What does the next training paradigm look like?Dwarkesh Patel
Created: 8/12/2026, 6:40:37 PM

My Notes

Loading notes...