Provider record · updated 2026-09-10
Together AI (GPU Clusters) — GPU pricing & deliverability
6 tracked offerings · grade mix A:0 · B:2 · C:0 · D:4 · U:0. GPU Clusters page publishes 'high-speed InfiniBand' interconnect for all clusters, managed Kubernetes or Slurm-on-Kubernetes orchestration, on-demand with no commitment, and reserved terms up to 6 months with upfront payment. Capacity described as 25+ cities: USA (600MW), Europe (150MW+), Asia and Middle East. Also listed reserved-only with contact-sales pricing: HGX B300 (270 GB), GB200 NVL72 (186 GB, 512 to 1,000+ GPUs), GB300 NVL72 (288 GB) — not recorded as offerings since no public price. No A100, L40S, or MI300X on the clusters page. InfiniBand generation/bandwidth (e.g. NDR 3.2Tbps) not specified on fetched page.
Tracked offerings
| GPU | SKU | $/GPU-hr | Terms | Grade | Fabric (as published) | Src |
|---|---|---|---|---|---|---|
| H100 | H100 (reserved) | $3.19 | reserved | D | High-speed InfiniBand | ↗ |
| H200 | H200 (reserved) | $3.99 | reserved | D | High-speed InfiniBand | ↗ |
| H100 | H100 | $3.99 | on-demand | D | High-speed InfiniBand | ↗ |
| H200 | H200 | $5.99 | on-demand | D | High-speed InfiniBand | ↗ |
| B200 | HGX B200 (reserved) | $6.79 | reserved | B | High-speed InfiniBand | ↗ |
| B200 | HGX B200 | $8.19 | on-demand | B | High-speed InfiniBand | ↗ |
What Together AI (GPU Clusters) does not publish
- gpus_per_node
- intra_node NVLink (not explicitly stated for B200/H200/H100)
- InfiniBand generation/bandwidth
- form_factor
- exact reserved minimum term
- regions per SKU
Unpublished fields are recorded as unknown and grade conservatively. Sellers improve their grades by disclosing — that is the point.
Weekly context on every provider
The Delivered Compute Report tracks how these prices and grades move.
Subscribe free