← research bank

NVIDIA's Rent Compresses Through Software, Not Silicon

nvidia-rent-runs-through-cuda · conviction medium · status open · horizon 2027 · as of 2026-08-01

The desk's generative scan rates nvidia-accelerators at the maximum contested score (incentive EXTREME x capacity HIGH = 12) with a measured 64% operating margin as the rent at stake. Consensus reads the entrant as silicon — hyperscaler ASICs built with Broadcom and Marvell. The variant: competitive silicon is necessary and not sufficient, because the binding constraint on substitution is whether frontier workloads RUN on it without hand-tuning. Rent compresses on the portability clock, not the tape-out clock — which means the tell is in framework and compiler releases, not in chip announcements.
Robust to undisclosed shares. 2 derived inputs under this thesis; redrawing every supply weight the industry does not publish moves none of them by more than 25%. Computed from evidence at most 17 days old (oldest input: alibaba).

Exhibits

Exhibit 1Relative performance, indexed to 100How the names in this thesis have traded against SOXX.
54245436MRVL 283SOXX 225AVGO 141NVDA 12312mo, indexed to 100 at start · dashed = SOXX benchmark

Series available as data/nvidia-rent-runs-through-cuda.csv

Exhibit 2Who pays CoWoS advanced-packaging capacity, and who keeps the moneyCapturers average 47.1% operating margin against payers' 44.6% — the owners of the scarce thing capture the rent, as expected.
Taiwan Semiconductor Manufac56.1%Analog Devices, Inc.38.1%NVIDIA Corporation64.0%SK Hynix58.6%Broadcom Inc.44.2%Advanced Micro Devices11.8%

Green/blue = model marks it as CAPTURING the rent (unbound and supplies the scarce good); faded = PAYING it (bound severe or moderate). Operating margin, live.

Exhibit 3What the conviction is actually made ofEach premise and the number it composes to. A conjunction of plausible premises is far weaker than any of them.
The rent is real, large, and measured — this is not a story about a company that might be profitableNVIDIA Corporation99.0%NVIDIA Corporation — operating margin 64.0% †100.0%Custom AI ASIC (XPU)90.0%COMPOSED (and)89.1%

† 1 premise marked supporting — shown and arguable, but the conclusion does not depend on them, so they are not multiplied into the composed figure. Citing a filed figure should not cost conviction.

89% if the 2 gates are independent, 90% if they move together. They are claims about one industry, so the truth is between and nobody can say where. Treat this as an ordering device rather than a calibrated probability — the ranking of premises is the information, not the level.

Weakest link: Custom AI ASIC (XPU) at 0.90 — Custom accelerators exist and are deployed at scale — Google TPU is multi-generation. The entrant is real, which is what makes the node contested rath

Substitution is gated by software portability, not by silicon availabilityCUDA and the portability question85.0%Inference serving stack (vLLM / TensorRT-LL…85.0%CUDA and the portability… — category softwa… †90.0%Incentive × Capacity — the indigenization /…60.0%COMPOSED (and)43.3%

† 1 premise marked supporting — shown and arguable, but the conclusion does not depend on them, so they are not multiplied into the composed figure. Citing a filed figure should not cost conviction.

43% if the 3 gates are independent, 60% if they move together. They are claims about one industry, so the truth is between and nobody can say where. Treat this as an ordering device rather than a calibrated probability — the ranking of premises is the information, not the level.

Weakest link: Incentive × Capacity — the indigenization / margin-compression generator at 0.60 — The desk's own prior, explicitly medium-conviction and NOT backtested. This thesis DISPUTES its capacity scoring for this node — capacity is rated HIG

Therefore the rent retreats toward the frontier rather than collapsing, and the observable is a portability milestone rather than NVIDIA Corporation — operating margin 64.0%100.0%CUDA and the portability question85.0%NVIDIA Corporation — demand pull at least 1070.4%Custom AI ASIC (XPU)90.0%COMPOSED (and)53.9%

54% if the 4 gates are independent, 70% if they move together. They are claims about one industry, so the truth is between and nobody can say where. Treat this as an ordering device rather than a calibrated probability — the ranking of premises is the information, not the level.

Weakest link: NVIDIA Corporation — demand pull at least 10 at 0.70 — REACTIVE. NVIDIA customer-weighted growth — currently 19.1 — is the demand leg. If the hyperscalers buying accelerators stop growing, the rent erodes

The variant

Consensus

Hyperscalers have overwhelming incentive to escape a supplier earning 64% operating margins on their largest capex line, and now have credible silicon: Google TPU is multi-generation, and Broadcom and Marvell are building custom accelerators for the other hyperscalers. Custom share therefore rises, merchant pricing power erodes, and NVIDIA's margin mean-reverts toward a normal semiconductor level.

Variant

The silicon is the easy half and it is already largely solved. What is not solved is that frontier training and serving stacks are written against one vendor's kernels first, so a competing accelerator inherits a porting cost measured in engineer-months per workload, paid again at every model architecture change. That cost is the actual moat, and it is invisible in any comparison of chip specifications. Consequently: custom silicon takes share fastest in STABLE, high-volume, internally-owned workloads (recommendation serving, ads ranking, first-party inference) and slowest at the frontier, where architectures move faster than the porting cost can be amortised. NVIDIA's rent does not collapse — it RETREATS toward the frontier, and the margin path depends on how fast the frontier itself commoditises.

Differentiator

Everyone models the substitution as a silicon race and dates it by tape-outs and foundry slots. The desk's own prior says a contested node needs incentive AND capacity, and scores capacity HIGH from the silicon side alone. This thesis argues capacity is gated by a SOFTWARE variable the ontology now holds explicitly (cuda-lock-in) and that nobody prices — so the observable that matters is compiler and framework portability milestones, not chip launches.

Indicators

Falsifiers

Open questions

Reasoning chain

The rent is real, large, and measured — this is not a story about a company that might be profitable VALID
premises

Both halves of a contested node are established: an extreme rent and a real entrant. The dispute is entirely about the RATE of substitution.

Substitution is gated by software portability, not by silicon availability VALID
premises

The weakest premise is deliberately the desk's own prior, because this conclusion is partly an argument against how that prior scored this node.

Therefore the rent retreats toward the frontier rather than collapsing, and the observable is a portability milestone rather than a chip launch VALID
premises

Composes lowest of the three, correctly: the conclusion stacks a structural software claim on a demand condition, and either can fail independently.

Sources

Write-up

Pre-filled skeleton: nvidia-rent-runs-through-cuda.md