Modern reasoning AI consumes enormous numbers of tokens internally — essentially thinking out loud before responding. That internal computation is inference, and speed directly translates to more reasoning cycles per dollar. Run Cerebras for 24 hours and you get the equivalent of weeks of AI thinking.