Dwarkesh Podcast

Podbit · Dwarkesh Podcast

Eric Jang – Building AlphaGo from scratch

Explore episode May 15, 2026

Where this was said

Alternative RL approaches

At 1:28:00 · chapter starts 1:25:38

Eric explains neural fictitious self-play — training best-response policies against fixed opponents and distilling them — as an MCTS substitute for games without tractable tree search.

Similar podbits