An individual user running batch size 1 may suffer more than 2 orders of magnitude worse compute efficiency compared to a large organization efficiently serving personalized weights.
Snapshot · Dwarkesh Podcast
An individual user running batch size 1 may suffer more than 2 orders of magnitude worse compute efficiency compared to a large organization efficiently serving personalized weights.
Where this was said
At 8:08 · chapter starts 7:32
The final prediction dives into the technical economics of inference at scale. Patel explains that serving a set of weights efficiently requires many concurrent sequences to be decoded against it simultaneously — a concept he explored in depth with Ryder Pope in a prior episode [1] — Dwarkesh Patel "Serving personalized AI weights efficiently requires batching thousands of sequences simultaneously — the optimal batch size for a sparse m…" 07:32 . Back-of-the-envelope estimates suggest the optimal batch size for a sparse model like DeepSeek V3 exceeds 2,400 concurrent sequences; falling short of this means leaving compute on the table. A large enterprise with thousands of employees and agents running diverse tasks can hit that batch size, efficiently utilizing its personalized weight fork. An individual user, by contrast, runs at batch size 1 — potentially suffering more than two orders of magnitude worse efficiency. The economic conclusion is sharp: personalized AI weights are a corporate-scale technology. The costs and efficiencies of the continual learning era will flow disproportionately to large organizations, not individuals. Patel closes by acknowledging that the most important consequences of continual learning are probably the ones hardest to anticipate — but the eight above seem clear enough from here.
Serving personalized AI weights efficiently requires batching thousands of sequences simultaneously — the optimal batch size for a sparse model like DeepSeek V3 exceeds 2,400. Individual users running batch size 1 face more than 100x worse compute efficiency, meaning the economics of personalized AI strongly favor large organizations.
Ad-based monetization works well for game apps where users spend extended time in-session, as seen with Grid and Wordle.
Tool-focused apps like PuffCount are poor candidates for ad monetization because users don't stay in-session long enough.
A hard paywall is a screen that blocks all app features unless the user pays or starts a free trial — it cannot be dismissed.
Mobile apps are primarily monetized through either ads (best for games) or in-app purchases/subscriptions (best for tools).
According to the episode, YouTube outperforms every other social platform for building trust and driving SaaS conversions.
Vasco stated that the majority of his app's user base came directly from his YouTube channel.
SEO Bot features a 'Boost My Domain Rating' button that routes users directly to Listing Bot, an example of in-product cross-selling.
The founder's entire product portfolio is AI-related, making it easier to package products attractively for directories.
The founder attached their SaaS demo to the trending debate about whether AI coding is actually good enough to build a full SaaS product.
We use essential and analytics cookies to run Vuci. To understand how the site is used: Privacy Policy.
Install Vuci on your phone
Add it to your home screen for a faster, app-like experience.
Install Vuci on your phone
Tap the Share button, then “Add to Home Screen”.
A new version is available
Reload to get the latest Vuci.