forecast

Old RL Ideas Should Combine With Large Models

Many ideas from earlier deep reinforcement learning work like DQN and AlphaGo are coming back into fashion and should be recombined with the new advances in large multimodal models, with significant potential in merging older and newer ideas.

forecastpending

Speaker

Demis Hassabis

Evidence Quote

I do actually think a lot of those ideas need to come back in again

Source

Demis Hassabis — Scaling, superhuman AIs, AlphaZero atop LLMs, AlphaFoldDwarkesh Patel Podcast
Created: 6/13/2026, 3:35:13 AM

My Notes

Loading notes...