Why would anyone pay SpaceX or Nebius $30-$50B/year/1GW of compute?
Because OpenAI and Anthropic can generate $100B+ per gigawatt per year selling inference API!
Here's how that's possible, and why it's sustainable (save this)
First, what are they actually selling? Every ChatGPT answer, every Cursor autocomplete, every enterprise copilot runs on "inference", the model generating tokens. The labs sell those tokens through subscriptions or an API, metered like electricity
@SemiAnalysis showed their model: run one gigawatt of Nvidia GB300 compute selling tokens at posted API prices and it generates over $100 BILLION a year of revenue. That same gigawatt costs roughly $12B-$50B/year to rent out
A 2-8x spread between what compute costs and what intelligence sells for
So why can they charge that much? Because the customer isn't comparing token prices to compute prices. They're comparing tokens to LABOR
A few dollars of tokens replaces work that costs hundreds of dollars an hour. Legal review, sales ops, financial close, code. That's why enterprise agent adoption is up 20x to 108x across job functions in just five months. At today's prices the buyer's ROI already fantastic
Why it's sustainable:
1. Demand compounds faster than prices fall. Token prices drop constantly, but agents burn dramatically more tokens per task than chatbots ever did, and every job function is adopting at once. Falling price x exploding volume = growing revenue
2. Supply is rationed. A handful of frontier labs, and none of them have enough compute. When you're capacity constrained you serve the highest-value demand first and pricing holds
3. The buyers keep paying UP, not down. Microsoft sells this same inference through Azure and Copilot. Nebius just disclosed its first deal at $40-50M per megawatt, the top of its own range. Nobody negotiates prices higher on a product that's about to be oversupplied
This is why the "AI capex bubble" framing keeps missing. The $100B at the top of the stack is what pays the $30-50B compute deals, which pay the datacenters, the chips, the memory, the power. The most profitable product in tech is funding everything below it
And you don't need to own the private labs to win. Every dollar of inference revenue flows down through the infra stack, and that's exactly where I'm positioned (compute, memory, power)
If this was helpful, my company provides a service where 5 top-tier analysts share their market analysis and real-time portfolios so you can see exactly how we're positioned across this stack. It's inside Milk Road PRO and just $1 to try (insane price just to check it out). Learn more here:
Follow me
@kylereidhead for more insights on AI, robotics and markets!