
Show notes
Eric Jang walks through how to build AlphaGo from scratch, but with modern AI tools. Sometimes you understand the future better by stepping backward. AlphaGo is still the cleanest worked example of the primitives of intelligence: search, learning from experience, and self-play. You have to go back to 2017 to get insight into how the more general AIs of the future might learn.
Highlighted moments
In AlphaGo, you don't train the policy network to imitate the MCTS action. You train it to imitate the MCTS distribution.
Transcript
Transcript not available for this episode yet.
More from Dwarkesh Podcast

Adam Brown – A deep but accessible introduction to general relativity
Jul 10, 20261h 38m

Grant Sanderson – AI and the future of math
Jun 30, 20261h 33m

The next big breakthrough will be AIs learning on the job
Jun 26, 202619 min

The data black hole at the center of AI
Jun 19, 202611 min

Ada Palmer – Machiavelli is the most misunderstood thinker of all time
Jun 16, 20262h 8m