Catalog

Every value the calculator reads, with the kind of evidence behind it and the source it came from. Researched to 2026-09-13.

Showing 21 hardware

H100 SXM

NVIDIA · accelerator

Accelerators
1 × H100 SXM
Memory
80 GB
Bandwidth
3.4 TB/s
Accelerator interconnect
0.9 TB/s
Power
0.7 kW
Acquisition
not published
Deployment multiplier
×1

Unpriced accelerator; use an explicit scenario acquisition price.

Official specifications.

H200 8-GPU system

NVIDIA · node

Accelerators
8 × H200
Memory
1,128 GB
Bandwidth
38.4 TB/s
Power
10.2 kW
Acquisition
$370,000
Reported range
$320,000 – $420,000
Deployment multiplier
×1.2

Independent typical 8-GPU estimate; planning power assumption.

Memory/bandwidth use official H200 values; the 8-GPU price range is independent research and power is a planning assumption.

DGX B200

NVIDIA · node

Accelerators
8 × B200
Memory
1,440 GB
Bandwidth
64 TB/s
Accelerator interconnect
14.4 TB/s
Power
14.3 kW
Acquisition
not published
Deployment multiplier
×1.2

Full-node acquisition price not published; planning assumption required.

Official node specifications; price intentionally missing.

DGX B300

NVIDIA · node

Accelerators
8 × B300
Memory
2,304 GB
Bandwidth
64 TB/s
Accelerator interconnect
14.4 TB/s
Power
14.5 kW
Acquisition
$350,000
Reported range
$300,000 – $500,000
Deployment multiplier
×1.2

No NVIDIA MSRP exists. Independent 2026 reports cluster at $300k-350k baseline with resellers quoting up to $500k; $350k is the recorded planning point.

CORRECTED at the 2026-08-01 refresh. This entry previously carried $550,000, a figure that re-verification could not trace to any source; current independent reporting puts an 8-GPU B300 node at $300k-350k baseline and up to $400-500k from resellers, so the point value is now $350k with the spread recorded. Memory (2,304 GB), 64 TB/s aggregate HBM bandwidth and 14.5 kW are all exact matches to NVIDIA's own DGX B300 User Guide. Note that the 14.4 TB/s figure circulating in secondary write-ups is the NVLink switch fabric, not memory bandwidth. Deployment multiplier is an editable planning assumption.

GB200 NVL72

NVIDIA · rack

Accelerators
72 × GB200
Memory
13,400 GB
Bandwidth
576 TB/s
Accelerator interconnect
130 TB/s
Power
120 kW
Acquisition
not published
Deployment multiplier
×1.25

Rack price not published; planning assumption required.

CORRECTED at the 2026-08-01 refresh: power is now NVIDIA's own approximately 120 kW from the DGX GB200 NVL72 User Guide, replacing the 132 kW previously recorded. CONFLICT preserved: infrastructure partners co-designing reference racks with NVIDIA (Vertiv, Schneider Electric, HPE) provision for up to 132 kW nominal with far higher transient peaks, so 132 kW is the right number for facility sizing and 120 kW for the rack itself. Memory and fabric are official; the 13.4 TB rack aggregate is NVIDIA's own published figure and deliberately is not 72 x a per-GPU value.

GB300 NVL72

NVIDIA · rack

Accelerators
72 × GB300
Memory
20,736 GB
Bandwidth
576 TB/s
Accelerator interconnect
130 TB/s
Power
135 kW
Acquisition
$3,500,000
Reported range
$3,000,000 – $4,000,000
Deployment multiplier
×1.25

No MSRP. Press and analyst reporting spans $3.0-4.0M per rack; $3.5M is the recorded midpoint.

CORRECTED at the 2026-08-01 refresh. The previous $3.0M simply reused the GB200 estimate, while reporting consistently puts GB300 at a premium over its predecessor ($3.0-4.0M, with one analyst note derived from a large customer order suggesting $3.7-4.0M); the midpoint is now recorded with the spread. Power moved from 132 kW to 135 kW: NVIDIA publishes no rack figure for GB300 at all, OEM reference documentation from HPE and Lenovo converges on 132-135 kW nominal, and other trackers report 140-142 kW — treat 132-142 kW as the honest range. Memory is derived as 72 x 288 GB; NVIDIA's own page rounds it to 20 TB.

Vera Rubin NVL72 (VR200)

announced

NVIDIA · rack

Accelerators
72 × Rubin (VR200)
Memory
20,736 GB
Bandwidth
1,584 TB/s
Accelerator interconnect
260 TB/s
Power
190 kW
Acquisition
not published
Deployment multiplier
×1.25
Available from
2026-07-01

No vendor price published. An analyst estimate of $3.5-4.0M per rack circulates but is not NVIDIA's figure; supply your own quote.

NAMING: NVIDIA announced this rack as "NVL144" at GTC 2025 counting compute dies, then as "NVL72" at CES 2026 counting GPU packages. They are the same rack, not two products. Memory (288 GB HBM4 x 72), bandwidth and NVLink 6 fabric are vendor figures. The 190 kW power draw is NOT vendor-published — it is the low end of a supply-chain analyst range of roughly 190 kW (Max-Q) to 230 kW (Max-P), so treat facility power here as an editable planning assumption. Declared in full production at CES 2026 with volume availability in the second half of 2026. DOWNGRADED 2026-08-29: this record was graded official-claim at medium confidence while citing no primary source at all — both sourceIds are independent-research, ServeTheHome and SemiAnalysis, so nothing here was read from NVIDIA. The gb300-nvl72 record carries the identical caveat about an unpublished rack power figure, cites one primary source more than this one, and is graded estimated at low confidence; the entry with the weaker evidence held the stronger grade. Nothing about the figures changed — 288 GB HBM4 x 72 and the NVLink 6 fabric are still what NVIDIA announced at CES 2026 and what those two outlets reported — only the claim this catalog makes about how it knows them.

Instinct MI455X Helios rack

announced

AMD · rack

Accelerators
72 × MI455X
Memory
31,104 GB
Bandwidth
1,670 TB/s
Accelerator interconnect
260 TB/s
Power
190 kW
Acquisition
not published
Deployment multiplier
×1.25
Available from
2026-09-01

No vendor price published. An analyst estimate of $5.0-5.5M per rack circulates but is not AMD's figure; supply your own quote.

NAMING: AMD's own press releases call this generation both "MI450 Series" (the OpenAI and Anthropic partnership announcements) and "MI400 Series" with the MI455X as flagship (the July 2026 launch). Same hardware. 432 GB HBM4 per accelerator across 72 accelerators is a vendor figure, as is the 260 TB/s UALink-over-Ethernet fabric. CONFLICT preserved on power: coverage of the same launch event cites both roughly 140 kW and 225-245 kW of bus-bar capacity under load; AMD has published no single authoritative rack figure, so 190 kW is recorded as a midpoint planning assumption and should be edited to whatever a vendor quote states. Declared in full production 2026-07-23 with shipments from Q3 2026.

Instinct MI300X 8-GPU platform

AMD · node

Accelerators
8 × MI300X
Memory
1,536 GB
Bandwidth
42.4 TB/s
Power
8.5 kW
Acquisition
not published
Deployment multiplier
×1.2

Full-node power planning value; acquisition price not published.

Memory and bandwidth are official. The data sheet's GPU-only 6 kW is converted to an editable 8.5 kW full-node planning value for TCO.

Instinct MI325X 8-GPU platform

AMD · node

Accelerators
8 × MI325X
Memory
2,048 GB
Bandwidth
48 TB/s
Power
10.5 kW
Acquisition
not published
Deployment multiplier
×1.2

Full-node power planning value; acquisition price not published.

Memory and bandwidth are official. The data sheet's GPU-only 8 kW is converted to an editable 10.5 kW full-node planning value for TCO.

Instinct MI355X 8-GPU platform

AMD · node

Accelerators
8 × MI355X
Memory
2,304 GB
Bandwidth
64 TB/s
Power
14 kW
Acquisition
$325,000
Deployment multiplier
×1.25

Node acquisition and 14 kW planning assumptions; not MSRP.

Official GPU specifications; node price/power are editable planning assumptions.

Gaudi 3 8-OAM node

Intel · node

Accelerators
8 × Gaudi 3
Memory
1,024 GB
Bandwidth
29.6 TB/s
Accelerator interconnect
4.2 TB/s
Power
9.5 kW
Acquisition
not published
Deployment multiplier
×1.2
Available from
2025-05-19

US$125,000 list for the 8-OAM baseboard incl. networking (Computex 2024); host system excluded, full-node price not published.

128 GB HBM2e and 900 W per accelerator are official; the 9.5 kW full-node figure converts the 7.2 kW accelerator-only TDP into an editable planning value (same convention as the AMD nodes).

RTX PRO 6000 Blackwell

NVIDIA · accelerator

Accelerators
1 × RTX PRO 6000 Blackwell Workstation Edition (GB202)
Memory
96 GB
Bandwidth
1.8 TB/s
Power
0.6 kW
Acquisition
$16,000
Reported range
$8,000 – $16,000
Deployment multiplier
×1

NVIDIA publishes no list price. $16,000 is press reporting of the September 2026 street price; pre-orders opened around $8,000–8,565 in March 2025, so the range spans that rise rather than a discount.

96 GB GDDR7 ECC, 1,792 GB/s and 600 W are NVIDIA's own specifications; the acquisition figure is not, and is recorded as a range because no vendor price exists to check it against. Roughly a doubling since launch, attributed by the same reporting to a GDDR7 shortage. Carries no NVLink: the datasheet lists PCIe 5.0 x16 as the whole of its system interface, and the Server Edition's product brief states 'NVIDIA NVLink: Not supported' outright for the same silicon. NVLink left this product line two generations ago — the RTX A6000 had a bridge for two cards at 112 GB/s, the RTX 6000 Ada that replaced it did not — so a second card in the same box talks over the host bus at 128 GB/s, not over a fabric.

RTX PRO 6000 Blackwell Max-Q

NVIDIA · accelerator

Accelerators
1 × RTX PRO 6000 Blackwell Max-Q Workstation Edition (GB202)
Memory
96 GB
Bandwidth
1.8 TB/s
Power
0.3 kW
Acquisition
not published
Reported range
$7,673 – $8,900
Deployment multiplier
×1

No vendor list price. The range is retailer and marketplace quotes gathered by search rather than read off a single page, so no midpoint is asserted.

Same 96 GB and the same 1,792 GB/s as the full-power Workstation Edition at half the board power: a power-capped bin for slim chassis, not a cut-down memory system. Compute throughput at 300 W is lower, and this catalog holds no measurement of by how much. Carries no NVLink: the datasheet lists PCIe 5.0 x16 as the whole of its system interface, and the Server Edition's product brief states 'NVIDIA NVLink: Not supported' outright for the same silicon. NVLink left this product line two generations ago — the RTX A6000 had a bridge for two cards at 112 GB/s, the RTX 6000 Ada that replaced it did not — so a second card in the same box talks over the host bus at 128 GB/s, not over a fabric.

4× RTX PRO 6000 Blackwell Max-Q

NVIDIA · node

Accelerators
4 × RTX PRO 6000 Blackwell Max-Q Workstation Edition (GB202)
Memory
384 GB
Bandwidth
7.2 TB/s
Accelerator interconnect
0.128 TB/s
Power
1.2 kW
Acquisition
not published
Reported range
$30,692 – $35,600
Deployment multiplier
×1.25

Four times the single-card range, which is itself retailer quotes rather than a vendor price. Nobody sells this as a product, so there is no build price to quote. The 1.25 multiplier stands for the host workstation the four cards need and is an assumption, not a configured quote.

Four is the count NVIDIA itself names: the Max-Q datasheet says "Scale up to four RTX PRO 6000 Max-Q GPUs", and the 600 W Workstation Edition datasheet makes no such claim, which is why this build is the 300 W part. Memory and bandwidth are four times the card’s own published figures. Power is board power for the four cards and excludes the host, matching how the single-card entries are recorded. The H100 page is cited for one thing only, and a review was right that nothing said so: NVIDIA prints "NVLink: 900GB/s | PCIe Gen5: 128GB/s" in one spec table, which fixes both the figure and the convention it is on, since the 900 is the bidirectional total this catalog already stores for H100 SXM. The A6000 page dates the loss of NVLink from the line.

RTX PRO 6000 Blackwell Server Edition

NVIDIA · accelerator

Accelerators
1 × RTX PRO 6000 Blackwell Server Edition (GB202, passively cooled)
Memory
96 GB
Bandwidth
1.6 TB/s
Power
0.6 kW
Acquisition
not published
Deployment multiplier
×1.2

Sold through server OEMs only; no public price.

The rack card carries the same 96 GB GDDR7 as the two workstation editions but publishes 1,597 GB/s against their 1,792 GB/s — a lower memory clock, stated on both NVIDIA's own page and an OEM product guide, so it is not a typo on one page. Decode is bound by that figure, which makes the server part the slowest of the three despite the datacenter label. Carries no NVLink: the datasheet lists PCIe 5.0 x16 as the whole of its system interface, and the Server Edition's product brief states 'NVIDIA NVLink: Not supported' outright for the same silicon. NVLink left this product line two generations ago — the RTX A6000 had a bridge for two cards at 112 GB/s, the RTX 6000 Ada that replaced it did not — so a second card in the same box talks over the host bus at 128 GB/s, not over a fabric.

RTX 6000 Ada Generation

NVIDIA · accelerator

Accelerators
1 × RTX 6000 Ada Generation (AD102)
Memory
48 GB
Bandwidth
0.96 TB/s
Power
0.3 kW
Acquisition
$6,800
Deployment multiplier
×1

Reseller listing; NVIDIA publishes no list price for this card either.

Previous generation and still sold new. It is here because "the 96 GB one or the 48 GB one" describes two generations rather than two configurations of one card: 96 GB is Blackwell, 48 GB is Ada, and the Ada part also has barely half the memory bandwidth.

GeForce RTX 5090

NVIDIA · accelerator

Accelerators
1 × GeForce RTX 5090 (GB202, consumer)
Memory
32 GB
Bandwidth
1.8 TB/s
Power
0.6 kW
Acquisition
not published
Reported range
$1,999 – $4,000
Deployment multiplier
×1

$1,999 is the MSRP; the upper bound is the September 2026 street price, which has stayed well above MSRP through the same memory shortage that moved the RTX PRO 6000.

Same GB202 die and the same 1,792 GB/s as the RTX PRO 6000, with a third of the memory. That pairing is why it is listed: on a consumer card it is capacity that runs out first, not bandwidth. NVIDIA's own page rendered neither price nor bandwidth in this session's fetch, so both are corroborated rather than read off the vendor.

DGX Spark

NVIDIA · node

Accelerators
1 × GB10 Grace Blackwell Superchip (unified CPU+GPU package)
Memory
128 GB
Bandwidth
0.273 TB/s
Power
0.2 kW
Acquisition
$4,699
Deployment multiplier
×1

NVIDIA published price, September 2026. It opened at $3,999.

128 GB of unified LPDDR5x buys capacity no discrete desktop card offers, at 273 GB/s — about a seventh of an RTX PRO 6000. Decode is memory-bound, so this box fits models it cannot serve quickly, and the catalog holds no throughput measurement for it. Power is the 240 W supply rating; the GB10's own TDP is 140 W.

Radeon AI PRO R9700

AMD · accelerator

Accelerators
1 × Radeon AI PRO R9700 (Navi 48, RDNA 4)
Memory
32 GB
Bandwidth
0.64 TB/s
Power
0.3 kW
Acquisition
$1,299
Deployment multiplier
×1

AMD retail price at launch; partner cards run $1,244 to $1,329.

32 GB GDDR6 at 640 GB/s and 300 W, from AMD's own product page. The price is press reporting of the launch figure, not a line on that page.

Mac Studio (M5 Max, 128 GB)

Apple · node

Accelerators
1 × Apple M5 Max (unified CPU+GPU+NPU, 40-core GPU bin)
Memory
128 GB
Bandwidth
0.614 TB/s
Power
0.5 kW
Acquisition
not published
Deployment multiplier
×1
Available from
2026-08-25

$2,499 buys the base machine, which is 36 GB at 460 GB/s — not this configuration. Apple's price for the 128 GB, 40-core-GPU build was not read off a vendor page in this session, so no acquisition figure is recorded.

Apple publishes two M5 Max bandwidths, 460 GB/s for the 32-core GPU and 614 GB/s for the 40-core; this entry is the higher bin, so a cheaper Mac Studio is also a slower one. Every throughput profile in this catalog was measured on a CUDA or ROCm serving stack and none of them transfers to Metal, so this system carries no curated rate and needs a manual one.

Inference Economics · static, local-first calculationCatalog 0.8.1 · cutoff 2026-09-13 · app v0.8.1 · build 1c759ca