forecast
Old RL Ideas Should Combine With Large Models
Many ideas from earlier deep reinforcement learning work like DQN and AlphaGo are coming back into fashion and should be recombined with the new advances in large multimodal models, with significant potential in merging older and newer ideas.
forecastpending
Speaker
Demis HassabisEvidence Quote
“I do actually think a lot of those ideas need to come back in again”
Source
Demis Hassabis — Scaling, superhuman AIs, AlphaZero atop LLMs, AlphaFold— Dwarkesh Patel PodcastCreated: 6/13/2026, 3:35:13 AM
My Notes
Loading notes...