Dwarkesh Podcast

Podbit · Dwarkesh Podcast

Eric Jang – Building AlphaGo from scratch

Explore episode May 15, 2026

Where this was said

Why doesn't MCTS work for LLMs

At 1:51:30 · chapter starts 1:41:09

Eric and Dwarkesh discuss why MCTS fails for LLM reasoning: language's vast token space violates the discrete finite-action assumption, and value estimation is much harder than in Go.

Technology
Why MCTS Doesn't Work for LLMs

Eric Jang – Building AlphaGo from scratch · May 15, 2026 Technology

Language has billions of possible next tokens — the PUCT exploration heuristic assumes you'll visit the same node multiple times, but an LLM will almost never generate the exact same token sequence twice. The discrete action assumption breaks.

Similar podbits

Science
The Placebo Effect Is Not Fake — It Just Belongs to the Story, Not the Drug

#403 ‒ Peptides: separating scientific promise from marketi… · Aug 10, 2026 Science

Peptides come loaded with a powerful narrative: regenerative, subcutaneous, targeted, cutting-edge. That narrative can genuinely move outcomes, especially for pain, energy, and recovery. The RCT's job is to quantify how much additional benefit comes from the molecule itself — above and beyond the compelling story surrounding it.