Reinforcement learning via backpropagation was a fundamental shift from supervised learning because it eliminated the need for paired labeled data, enabled exploration and mistakes (the only path to genuine intelligence), and mirrored how humans and animals learn through trial-and-error.
causalpending
Speaker
Mohamad GafaitEvidence Quote
“you no longer need the pairs of accurate labeled highly label data uh second is it allows exploration which I think is what most people don't understand is the only path to genuine intelligence is to explore and make mistakes”
Source
Ex-Google Exec Reveals The Shocking Truth About AI with Mo Gawdat | Bad Decisions Podcast #64— Bad Decisions StudioCreated: 8/11/2026, 7:05:25 AM
My Notes
Loading notes...