Provider record · updated 2026-09-10
Microsoft Azure — GPU pricing & deliverability
3 tracked offerings · grade mix A:2 · B:1 · C:0 · D:0 · U:0. Azure publishes per-GPU InfiniBand bandwidth on Microsoft Learn: ND H100 v5 and ND H200 v5 each give every GPU a dedicated 400 Gbps NVIDIA Quantum-2 ConnectX-7 InfiniBand link (3.2 Tbps per 8-GPU VM) with GPUDirect RDMA, scaling 'up to thousands of GPUs'. ND GB200 v6 (4 GPUs/VM) has 4x 400 Gb/s Quantum-2 CX7 IB per VM, NVL72 NVLink domains of 72 GPUs (18 VMs/rack, 28.8 Tb/s rack scale-out), and a stated InfiniBand fabric reach of up to 100,000 GPUs. Prices from the official Azure Retail Prices API (Consumption meters, Linux).
Tracked offerings
| GPU | SKU | $/GPU-hr | Terms | Grade | Fabric (as published) | Src |
|---|---|---|---|---|---|---|
| H200-SXM | Standard_ND96isr_H200_v5 | $10.60 | on-demand | B | InfiniBand NDR: dedicated 400 Gbps NVIDIA Quantum-2 CX7 InfiniBand per GPU, 3.2 Tb/s per VM, GPUDirect RDMA (Microsoft Learn) | ↗ |
| H100-SXM | Standard_ND96isr_H100_v5 | $12.29 | on-demand | A | InfiniBand NDR: 'dedicated, topology-agnostic 400 Gbps NVIDIA Quantum-2 CX7 InfiniBand connection' per GPU, 3.2 Tbps per VM, GPUDirect RDMA (Microsoft Learn) | ↗ |
| GB200-NVL72 | Standard_ND128isr_NDR_GB200_v6 | $27.04 | on-demand | A | InfiniBand NDR: 4x 400 Gb/s NVIDIA Quantum-2 CX7 InfiniBand connections per VM (one 400 Gb/s NIC per GPU); 28.8 Tb/s scale-out networking per 72-GPU rack (Microsoft Learn) | ↗ |
What Microsoft Azure does not publish
- Numeric maximum cluster size for ND H100 v5 / ND H200 v5 (Microsoft states only 'thousands of GPUs')
- ND96isr_H200_v5 price in eastus specifically (API listed eastus2/centralus and others, not eastus, at collection time)
Unpublished fields are recorded as unknown and grade conservatively. Sellers improve their grades by disclosing — that is the point.
Weekly context on every provider
The Delivered Compute Report tracks how these prices and grades move.
Subscribe free