Chips & Compute

NVIDIA Blackwell

NVIDIA · #15 most active of 18 in Chips & Compute

Compare →

NVL72 rack-scale coherent GPU domains, the default substrate for frontier training and inference.

Current version

GB200 NVL72

Entry cost

~$2-6 per GPU-hour on clouds

Changes / 30d

0

Capabilities

  • 72-GPU NVLink domain
  • FP4 inference
  • HBM3E 192GB
  • CUDA ecosystem

Pricing

~$2-6 per GPU-hour on clouds

Entry price over time

Not enough pricing history yet — we start charting from the second observation.

Geek mode

Models under the hood

  • Blackwell B200

    NVIDIA

    disclosed
  • Grace CPU

    NVIDIA

    disclosed

Context window

Public API

yes

Multi-model routing

no

Stack signals

  • CUDA
  • NVLink 5
  • TSMC 4NP
  • HBM3E

Shared model stack

Other tracked products running on the same foundation models — a quick read on how much of the catalog moves when one of these models changes.

green disclosed · amber inferred · grey unknown

Reported scale

Reference data. Each figure is whatever the source actually said — weekly users, downloads, revenue run-rate — with its own definition and date. These are not comparable between tools and are never used to rank anything.

No public usage figure on record for this product.

Also in Chips & Compute

All Chips & Compute

Wafer-scale inference and training

Version
WSE-3 / Inference
Cost
Free tier, then from $0.06 / M tokens
Model
Llama / Qwen / GPT-OSS

Inference-first TPU pods

Version
v7 Ironwood
Cost
from ~$1.2 per chip-hour
Model
TPU v7

Deterministic low-latency inference

Version
GroqCloud
Cost
Free tier, then from $0.05 / M tokens
Model
Llama / Kimi / GPT-OSS

High-memory GPU alternative

Version
MI355X
Cost
~$2-3 per GPU-hour on clouds
Model
CDNA 4

On-device inference across Mac, iPhone and iPad.

Version
M5
Cost
Bundled with hardware
Model
Apple Foundation Models

Cloud-native training silicon

Version
Trainium2
Cost
from ~$1.3 per accelerator-hour
Model
Trainium2

Change history

  • version

    GB200 NVL72 racks reach broad cloud availability

    HGX H200GB200 NVL72

    source
  • capability

    FP4 inference kernels land in TensorRT-LLM

    FP4 inference

    source