Syntax - Tasty Web Development Treats

Snapshot · Syntax - Tasty Web Development Treats

1021: We got addicted to an AI model we can't talk about

Explore episode Jul 15, 2026

Where this was said

Local vs Cloud AI Model Hosting

At 31:35 · chapter starts 30:27

Wes asks the inevitable: are AI inference costs going to spiral into thousands per employee per month? Dax has actual data to offer. At their 5x surge level, OpenCode's inference spend amounts to roughly 15% of their payroll — notable, but manageable for a tech company. And the structural trajectory is downward. His estimate: Anthropic and OpenAI are currently running approximately 90% margin on inference, excluding R&D. Breakeven is 10x cheaper than current prices. Training losses are separate — they don't factor into inference economics. For open-source models, OpenCode can already host at a 70% discount to cost even using GPU middlemen. Direct GPU ownership would approach those same 90% margins. The narrative that OpenAI and Anthropic are perpetually unprofitable confuses training investment with inference margins — these are very different line items on the P&L.

Similar snapshots