Dwarkesh Podcast

Snapshot · Dwarkesh Podcast

8 Predictions for the Era of Continual Learning

Explore episode Aug 7, 2026

Where this was said

Prediction 8: Inference Economics Will Strongly Favor Large Enterprises

At 8:08 · chapter starts 7:32

The final prediction dives into the technical economics of inference at scale. Patel explains that serving a set of weights efficiently requires many concurrent sequences to be decoded against it simultaneously — a concept he explored in depth with Ryder Pope in a prior episode. Back-of-the-envelope estimates suggest the optimal batch size for a sparse model like DeepSeek V3 exceeds 2,400 concurrent sequences; falling short of this means leaving compute on the table. A large enterprise with thousands of employees and agents running diverse tasks can hit that batch size, efficiently utilizing its personalized weight fork. An individual user, by contrast, runs at batch size 1 — potentially suffering more than two orders of magnitude worse efficiency. The economic conclusion is sharp: personalized AI weights are a corporate-scale technology. The costs and efficiencies of the continual learning era will flow disproportionately to large organizations, not individuals. Patel closes by acknowledging that the most important consequences of continual learning are probably the ones hardest to anticipate — but the eight above seem clear enough from here.

Similar snapshots