No positions in the markets we measure BY 07:30 ET
CCIR Compute Credit
Index Research
Monitor · Chip Economics

Chip Economics

Earning power per spec-sheet unit.

Analytical decomposition of the published daily rates: derived analysis.

01 Four Lenses on One Rate

Capacity $ / GB-hr spread 47%

what the chip can hold — the size of model and context it can host

.02.04 0y 3y 6y B300 B200 H200 H100 A100
Bandwidth $ / TB/s-hr spread 27%

how fast it moves data — the dimension inference workloads buy

.501.00 0y 3y 6y
Compute $ / PFLOP-hr spread 100%

how fast it does arithmetic — training speed

$2$4$6$8 0y 3y 6y
Power $ / kW-hr spread 32%

what it costs to keep lit — rent per watt against facility cost per watt

$2$4$6 0y 3y 6y
Blackwell Hopper Ampere same horizontal position = same generation in every panel · hover or focus any dot for all four readings

Neocloud (T2) guaranteed on-demand, all regions, generations on age (x, years) against each lens (y, from zero). Spread = (max − min) ÷ mean across the five generations. Example prints · 2026-09-28 · rates from the 2026-09-28 snapshot. Headline stat per published rule (operator-equal median). The compute lens is denominated in petaFLOPs (PFLOP, 10^15 floating-point operations per second), dense, in the BF16 number format. Blackwell chips can also run FP4, a 4-bit format that Hopper and Ampere chips cannot use. For inference that runs in FP4, this lens understates what newer chips deliver per dollar.

One observed number underlies every panel: the Neocloud guaranteed on-demand rate, in dollars per GPU-hour. Each panel divides that rate by one spec-sheet denominator. One of the four compresses the cross-section far more than the others: memory bandwidth, at a 27% spread. Power spreads 32%, capacity 47% and compute 100%. Power is a looser second.

GenerationAge$/GPU-hrn$/GB-hr$/TB/s-hr$/PFLOP-hr$/kW-hr
B300 ~0.9y $7.73 10 0.02861.0043.437.03
B200 ~1.5y $6.85 9 0.03810.8903.046.85
H200 ~2.2y $4.14 14 0.02940.8624.185.91
H100 SXM ~3.8y $3.56 21 0.04441.0613.595.08
A100 80GB ~5.8y $2.30 11 0.02871.1287.375.75

CRI-T2-{chip}-SXM-GTD-OD-ALL-{GB|BW|PF|KW} · derived CRI series carry the identifier of the underlying cell in the first seven segments; strip the lens suffix to reach the parent. $/GB-hr computed from vendor high-bandwidth memory (HBM) capacity (80 / 80 / 141 / 180 / 270 GB).

02 The Bandwidth Band

Re-denominated per unit of memory bandwidth, datacenter silicon from 0.9 to 5.8 years of age rents inside $0.862–1.128 per TB/s-hour, a 27% spread around a $0.989 mean. Raw, the same five generations span 3.4×. The band is the test: a spread that widens and stays wide is the falsifiable marker firing. Members are disclosed on every exhibit.

The compression is a like-for-like property. It holds across cells that share an operator segment, interruptibility grade, term, and region: the series named under each exhibit. Pooled cuts, marketplace minimums, or mixed interruptibility grades will not reproduce it: grade and tenor are different markets per unit of bandwidth, as everywhere else on this site.

Power is the next tightest lens, at 32%. Rent per watt sorts by cooling class, which §03 shows.

0.800.901.001.101.20 1.128 — max 0.989 — mean 0.862 — min 0y 2y 4y 6y age axis zoomed: $0.75–1.2, not to zero B300 1.00 B200 0.89 H200 0.86 H100 SXM 1.06 A100 80GB 1.13
$/TB/s-hr against generation age; band = min–max across qualifying members, mean line at $0.989. Denominator: vendor nameplate memory bandwidth. Example prints · 2026-09-28 · rates from the 2026-09-28 snapshot.

CRB-T2-BW-ALL-OD-ALL · band row fields: band_min · band_max · band_mean · band_spread_pct · n_chips · member_chips. Emission gate n_chips ≥ 3, all members publication-qualified. Research-grade at launch (Shadow → Provisional).

The two parity rules. The sharpest way to see what the band rules out is to price the oldest member both ways. If compute were the priced unit, every new generation would force the prior one down to its share of the throughput: against the H100 print of $3.56, a chip with 32% of the BF16 throughput reprices to $1.12. If bandwidth is the priced unit, the same H100 print implies $2.16: the A100 carries 61% of the bandwidth. Two rules, two implied rates:

Pricing ruleAssumes the market pays forImplied A100 rate
Compute parityBF16 throughput: 312 vs 989.5 TFLOPS$1.12
Bandwidth parityHBM bandwidth: 2.039 vs 3.35 TB/s$2.16

The observed A100 print is $2.30: 106% of bandwidth parity, 105% above compute parity, at 5.8 years of age. The market prices the A100 near its bandwidth share. That is the band's claim restated as a spread you can falsify: the day prior-generation prints migrate from the bandwidth-parity line toward the compute-parity line is the day the band breaks. The schedule to watch it on is the bandwidth frontier itself, which has moved more slowly than the compute frontier: the band's newest member entered at its predecessor's 7.7 TB/s.

03 Watt-Rent by Cooling Class

Per unit of nameplate power, the panel sorts by cooling class rather than by age. Air-cooled datacenter silicon rents at $4.49–5.75 per kW-hr (L40S, A100, H100). Liquid-class Blackwell rents at $6.85–7.03. H200, which runs in dense-air or liquid halls, sits between them at $5.91. Consumer cards on marketplaces, with no datacenter bundle, rent at $0.60–1.12. Within each class the prints sit close together, and each class sits on its own level. Watt-rent prices the hall a chip can occupy.

$0$1$2$3$4$5$6$7 $/kW-hr LIQUID-CLASS $6.85–7.03 AIR-COOLED DC $4.49–5.75 MKT / CONSUMER $0.60–1.12 B200 6.85 B300 7.03 H200 5.91 † A100 5.75 · ~5.8y H100 5.08 · ~3.8y L40S 4.49 3090 / 4090 / 5090 · $0.60–1.12 (Mkt)
Revenue per kW of nameplate thermal design power (TDP), by cooling class: gross rent per unit of power envelope, not margin. Neocloud (T2) guaranteed on-demand, all regions, as in §01. † H200 hosts in dense-air or liquid halls. Marketplace rows (Mkt) are Marketplace-segment prints, indicative. Rates from the 2026-09-28 snapshot.
table view — watt-rent panel
GenerationTDP$/GPU-hrn$/kW-hrCooling class
B3001.1 kW $7.73107.03liquid
B2001.0 kW $6.8596.85liquid
H200700 W $4.14145.91dense air / liquid
H100700 W $3.56215.08air
A100400 W $2.30115.75air
L40S350 W $1.5794.49air
3090 (Mkt)350 W$0.2140.60consumer
4090 (Mkt)450 W$0.4941.10consumer
5090 (Mkt)575 W$0.6541.12consumer

CRB-T2-KW-AIR-OD-ALL · CRB-T2-KW-LIQ-OD-ALL · cohort = cooling_class (air | liquid); consumer prints are Marketplace-segment context, not band members.

04 Economic Life: the Breakeven Altimeter

Watt-rent and cash operating cost are the same unit. A chip exits economic life when its watt-rent decays to the cash boundary beneath it. One log-scale strip therefore reads as an altimeter: today's rent bands above, the cash operating boundary below, and the gap between them is distance to shutdown. Air-cooled silicon today covers its cash operating cost ≈6–10×. The consumer band is the visible preview of late life, close above the boundary.

Cost-side basis. Chip TDP understates system draw. NVIDIA rates the 8×H100 DGX system at 10.2 kW maximum against 5.6 kW of chip TDP, a 1.82× system factor (CPUs, NVSwitch, fans). Every cost rung below is restated per chip-TDP-kW through that factor, so both sides of the altimeter share the watt-rent lens's denominator. The factor is a disclosed election.
$ / TDP-kW-hr · log scale $8$5$2$1$0.50$0.20$0.10 rent decays toward cost ≈6–10× cash coverage LIQUID-CLASS RENT $6.85–7.03 B300 7.03 · B200 6.85 · H200 5.91 (dense air or liquid) AIR-COOLED DC RENT $4.49–5.75 H100 5.08 · A100 5.75 · L40S 4.49 CONSUMER RENT $0.60–1.12 no datacenter bundle: the preview of late life SHUTDOWN BOUNDARY: CASH COST ≈$0.61–0.74 power + colo · sunk capex excluded · ≈$0.43–0.52 per H100-hr before staffing and operations COMPONENT: WHOLESALE COLO ≈$0.40–0.49 CBRE $160–196.25/kW-mo × 1.82 system draw COMPONENT: POWER ≈$0.21–0.25 EIA 8.91¢/kWh × PUE 1.3–1.54 × 1.82 system draw
Rent bands are CCIR prints (filled, member ticks at left edge). The cash boundary is a derivation from disclosed elections (tinted), its components external reference data (outlined). Whether power is metered inside or beside the colo rate varies by contract, so the two components are shown separately. All cost rungs per chip-TDP-kW via the 1.82× system factor. Sources: CBRE North America Data Center Trends (H2 2025: $196.25 per kW-month for 250–500 kW wholesale; H1 2026: $160–185 for 10-plus MW in Northern Virginia); U.S. Energy Information Administration (EIA) industrial price, trailing 12 months to July 2026; Uptime Institute (power usage effectiveness, PUE: new builds at 1.3 or lower, 2025 weighted average 1.54); NVIDIA DGX H100 user guide (10.2 kW max). Rates from the 2026-09-28 snapshot; cost inputs verified 2026-09-29.
The crossing identity. rent per watt = band level × (chip TB/s ÷ chip kW). Watt-rent is the chip's breakeven cash operating cost: economic life ends when watt-rent decays to cash cost per TDP-kW. Sunk capex is excluded, which is why this boundary governs whether an installed chip keeps running. Both sides are $/TDP-kW-hr. CCIR publishes the left side daily. The cost side (colo terms, power price, staffing, the system factor, any decay election) is the reader's.

Worked illustration: hypothetical; every input elected by the reader

Elect a cash operating boundary c = $0.68/TDP-kW-hr (inside the cash band above) and a decay election g = 21%/yr (the cross-generation pace; see the caveat below). Air-band watt-rent today ≈ $5.12/TDP-kW-hr.

years to crossing = ln(watt-rent ÷ c) ÷ g = ln(5.12 ÷ 0.68) ÷ 0.21 ≈ 10

This is arithmetic on today's band against two elected inputs: a band-implied rent set against an elected cash cost. It is not a projection of any rate, and CCIR publishes neither election.

The honest gap. The −21% per year slope is a cross-generation reading: how fast rent falls with age across five generations on one date (/research/gpu-age-curve). It is not the band's own decay through time, which is the parameter the crossing needs. This page does not measure that decay. The band is also an observed equilibrium, not a law. In the 2023 shortage, frontier chips were posted far above any spec-sheet line. The flatness carries a date, which is the reason to keep watching it.

05 Method & Caveats

Normalized values equal the published parent series value divided by the disclosed constant. The parent cell's pooling, stat, and gate decisions are inherited unchanged. Denominators are vendor nameplate constants from NVIDIA datasheets, linked in the table below. A correction is a changelog entry, not a series break. The parent cell's published headline statistic (the operator-equal median) carries through unchanged.

LensDenominatorA100 80GBH100 SXMH200B200B300Spec source
Capacity · $/GB-hrwhat the chip can hold: the size of model and context it can host HBM capacity (GB) 8080141180270 NVIDIA datasheets
Bandwidth · $/TB/s-hrhow fast it moves data: the dimension inference workloads buy memory bandwidth (TB/s) 2.0393.354.87.77.7 NVIDIA datasheets
Compute · $/PFLOP-hrhow fast it does arithmetic: training speed dense BF16 (PFLOPS) 0.310.990.992.252.25 NVIDIA datasheets · dense, not sparse
Power · $/kW-hrwhat it costs to keep lit: rent per watt against facility cost per watt TDP (kW) 0.40.70.71.01.1 NVIDIA datasheets

Denominators are variant-level: H100 SXM 3.35 TB/s and H100 PCIe 2.0 TB/s are different denominators. A chip without a disclosed denominator emits no row for that lens; nothing is imputed. Sources: A100 80GB · H100 SXM · H200 · B200 · B300: NVIDIA Blackwell Ultra datasheet, HGX B300 column (1,100 W TDP, 270 GB, 7.7 TB/s).

Lens admission rule

A lens is admitted when its denominator is a disclosed vendor nameplate constant, it answers a question a credit reader has, and it tells a distinct story. Considered and excluded on record: token throughput (measured, not nameplate), interconnect / fabric (a cluster property with no per-chip denominator; the tier axis prices the fabric), PUE-adjusted power (PUE is the reader's election), and FP8 / FP4 precision variants (a precision election; disclosure, not columns).

Blind spots.

  • Denominators are nameplate constants. They capture none of a deployment's realized capacity, interconnect domain, or measured throughput.
  • System draw exceeds chip TDP (1.82× for the DGX H100, per NVIDIA's rated 10.2 kW). TDP is the disclosed lens constant, and delivered-power cost comparisons apply the system factor.
  • PUE, facility cost, and any decay election are the reader's. CCIR publishes prices, not the elections.
  • The cross-section is one date across generations, not a cohort through time. §04 states the consequence.
  • All numerators are posted list asks, not transactions. Thin cells (n < 3) are indicative.

Full construction: /documents/methodology · underlying cells: /explorer · the dated study these lenses generalize: /research/gpu-age-curve.