
Show notes
Eric Jang walks through how to build AlphaGo from scratch, but with modern AI tools. Sometimes you understand the future better by stepping backward. AlphaGo is still the cleanest worked example of the primitives of intelligence: search, learning from experience, and self-play. You have to go back to 2017 to get insight into how the more general AIs of the future might learn.
Highlighted moments
In AlphaGo, you don't train the policy network to imitate the MCTS action. You train it to imitate the MCTS distribution.
Transcript
Transcript not available for this episode yet.
More from Dwarkesh Podcast

8 Predictions for the Era of Continual Learning
Aug 7, 20268 min

Why smarter AI models could drive up compute prices 10x
Aug 3, 202611 min

Adam Brown – A deep but accessible introduction to general relativity
Jul 10, 20261h 38m

Grant Sanderson – AI and the future of math
Jun 30, 20261h 33m

The next big breakthrough will be AIs learning on the job
Jun 26, 202619 min