All major AI labs are making a research bet that training AIs on millions of verifiable tasks across thousands of diverse RL environments will produce AGI, because such training creates general problem-solving agents capable of making progress on open-ended tasks for weeks despite errors, ambiguity, and mistakes.

factualpending

Speaker

Unidentified Speaker — What does the next training paradigm look like? [20p5-kQXF_Q]

Evidence Quote

So here's a big research bet that all the labs are making. They think that if we train AIs to accomplish millions of verifiable tasks across thousands of diverse RL environments, then we will have basically built AGI

Source

What does the next training paradigm look like?Dwarkesh Patel
Created: 8/12/2026, 6:40:37 PM

My Notes

Loading notes...