news

Nvidia's Megawatt Got 60% More Expensive and 30 Times More Productive in One Generation.

Nvidia disclosed both curves: a megawatt went $25m to $40m and got 30x more productive. Revenue per megawatt holds only if token prices fall less than 30x.

One Nvidia generation: 60% dearer per megawatt, 30x the output

Nvidia disclosures from the fiscal Q2 2027 call, with R40 arithmetic on the ratios

Per megawattFigure
Hopper$18M
Blackwell$25M
Vera Rubin$40M
Cost step, last generation+60%
Throughput step, same generation30×
Throughput per dollar of hardware18.8× better
Price fall that cancels it30×

The per-gigawatt series and the 30x throughput and 35x token-cost figures are Nvidia's own, disclosed on the call; the 30x and 35x are Vera Rubin against Grace Blackwell Ultra. The per-megawatt costs are what a buyer pays Nvidia for silicon, interconnect and networking - land, shell, substation, cooling and construction are not in them, and part of the rise is scope, since Vera Rubin's $40M spans a CPU line, two networking fabrics and a Groq LPU that Hopper's $18M did not. Throughput per dollar and the break-even price fall are ours, arithmetic on those disclosures.

Revenue per megawatt is a race between throughput and priceRevenue per megawatt, indexed to 1.0 before the generation step — R40 arithmeticRevenue per megawatt, indexed01.534.56Price -5× — Revenue per megawatt, indexed: 66Price -5×Price -10× — Revenue per megawatt, indexed: 33Price -10×Price -20× — Revenue per megawatt, indexed: 1.51.5Price -20×Price -30× — Revenue per megawatt, indexed: 11Price -30×Price -35× — Revenue per megawatt, indexed: 0.90.9Price -35×Throughput per megawatt rose 30x on Nvidia's disclosed figure. Revenue per megawatt is that multiplied by what a unit of outputsells for, so each column is 30 divided by the price fall in the label. At a 30x price fall revenue per megawatt is unchanged,which is the break-even; below it the operator captures the difference, above it the hardware improves while the business doesnot. Nvidia's own 35x token-cost improvement sits past the break-even, which is why volume growth rather than price is what hasto carry revenue per megawatt.

Buried in Nvidia's August call are the two numbers that decide whether an AI data centre is a good business, and they point in opposite directions.

The first is what a megawatt costs. Colette Kress walked the series generation by generation:

Since Hopper, our revenue opportunity has grown from roughly $18 billion per gigawatt to $25 billion with Blackwell, to $40 billion with Vera Rubin, which now spans Vera CPU, Rubin GPU, NVLink, InfiniBand or Ethernet, and Groq LPU.

The second is what that megawatt produces. From the same call, two sentences later:

Vera Rubin exemplifies this, delivering 30x higher throughput per megawatt and 35x lower token cost relative to Grace Blackwell Ultra.

Put them together and one generation of Nvidia hardware costs 60% more per megawatt and produces thirty times as much from it. That is an 18.8× improvement in throughput per dollar of hardware, in a single step, and it is the reason buyers keep paying more per megawatt rather than less.

But revenue per megawatt is not throughput per megawatt. It is throughput multiplied by what a unit of throughput sells for, and that second number is falling at least as fast as the first is rising. This piece is about what happens where those two curves meet, because that collision — not the price of the hardware — is what determines whether the megawatt pays for itself.

The points

What a megawatt costs, and what is in it

Generation Nvidia revenue per MW Step Cumulative
Hopper $18M 1.00×
Blackwell $25M +39% 1.39×
Vera Rubin $40M +60% 2.22×

Two things are worth separating here, because they get conflated constantly.

This is not the cost of a data centre. It is what the buyer pays Nvidia — silicon, interconnect and networking. Land, shell, substation, cooling, construction and operations sit on top and are not in any of these numbers.

And part of the rise is scope. Hopper's $18 million bought GPUs. Vera Rubin's $40 million buys a CPU line Nvidia did not previously sell at scale, two networking fabrics, and an LPU from the Groq partnership. Nvidia is capturing a larger share of the same build, so the per-megawatt figure rises even where the underlying components do not. Kress said as much on the call — the full-stack platform is "expanding our share of the data center TAM."

What a megawatt now produces

The 30× is the number that makes the 60% rational, and its unit matters. It is throughput per megawatt, not per chip or per rack. When the binding constraint on an AI build is power — and Nvidia spent much of the call saying supply and power are exactly what bind — throughput per megawatt is the only efficiency figure that converts directly into revenue.

The companion figure, 35× lower token cost, is the same improvement seen from the operator's side. It is what it costs them to produce a token, and it falls faster than the throughput rises because power efficiency improves alongside compute density.

So the operator gets thirty times more output from a megawatt they paid 1.6 times more for. On the cost side of their P&L this is unambiguously good, and it is why a $40 million megawatt is a better purchase than an $18 million one despite costing more than twice as much.

The collision: what actually happens to revenue per megawatt

Here is the part that is not on the slide. Revenue per megawatt is:

throughput per megawatt × price per unit of throughput

Nvidia has told us the first term rose 30×. Nobody has told us what the second is doing, and it is falling fast — that is the entire direction of travel in inference pricing, and it is the subject of our token-economics work. So the outcome for the operator is a race:

If the price of a unit of output falls Revenue per megawatt
6.0× higher
10× 3.0× higher
20× 1.5× higher
30× unchanged
35× 14% lower

Thirty times is the break-even. If prices fall by less than throughput rises, the operator captures the difference and revenue per megawatt goes up. If they fall by more, the hardware is getting better and the business is getting worse at the same time — which is a thing that can happen, and has happened in other commodity-compute markets.

Two observations sharpen it. First, Nvidia's disclosed 35× token-cost improvement is above the 30× throughput gain, which tells you the company itself expects unit costs — and therefore achievable prices — to fall faster than output rises. Second, prices in this market have been falling at rates that make 30× look modest: on ARC Prize's verified runs — the one place a solved reasoning task is priced the same way twice — the cost of the same 87.5% score fell from about $4,560 to $0.30 in twenty months.

That does not mean revenue per megawatt is falling. It means volume has to grow into the gap, which is precisely the Jevons argument — cheaper output opens workloads that were uneconomic before, and total tokens consumed rises faster than price per token falls. The 30× break-even is a clean way to state how much new demand the industry needs to find per generation just to stand still.

Who captures it

The same physical megawatt produces very different revenue depending on who owns the output, and the spread is larger than any of the hardware numbers above.

Owner of the megawatt Revenue per MW per year Months to cover a $40M Vera Rubin megawatt
Landlord renting capacity $4.47M 107
Blended landlord-plus-model-layer $5.68M 85
Frontier model provider (estimated) ~$50M 10

The frontier figure is an outside estimate, not a disclosure: roughly $50 million of revenue per megawatt for Anthropic in 2026, against a $10–15 million compute cost, published by Tomasz Tunguz and by Dylan Patel of SemiAnalysis. Anthropic's reported $10.9 billion of revenue and $559 million of operating profit are consistent with it. Nobody has filed it, and the row that uses it is directional.

The first two rows are ours, from our model of SpaceX — a capacity-driven model that treats it largely as a landlord: 55% of capacity leased at $2.03 million per megawatt-quarter, 15% reserved for the model layer Grok runs on. A landlord needs nine years of gross revenue to cover the hardware. A frontier lab needs ten months. That is the same silicon, drawing the same power, on the same site.

This is why the 30× matters more to a model provider than to whoever built the building. The throughput gain accrues to whoever sells the tokens. A landlord's rent does not rise 30× because the tenant's hardware got better; it rises with what the market will pay for a megawatt of hosted capacity, which is a different and much slower curve.

What would change the conclusion


The per-gigawatt series, the 30× throughput and 35× lower token cost for Vera Rubin against Grace Blackwell Ultra, and the description of what the platform spans are all disclosed by Nvidia on its fiscal Q2 2027 call and in the accompanying release, covered in our note on that print. The break-even table, the throughput-per-dollar figures and the payback months are ours, arithmetic on those disclosures. The roughly $50 million of revenue per megawatt for a frontier model provider and the $10–15 million compute cost beside it are estimates published by Tomasz Tunguz and by Dylan Patel of SemiAnalysis, not company disclosures. The landlord rows come from our SpaceX model, whose $2.03 million per megawatt-quarter price and split of capacity between leasing and the model layer are assumptions of ours; SpaceX is private and discloses no capacity, capital expenditure or supplier list.

Related

Stocks in this article