Buyer's guide · compiled 2026-09-10
The GPU contract gotcha list
Reserved GPU contracts are where the deliverability problem becomes a legal problem. Below: the clauses that burn buyers, compiled from published negotiation guides, provider rating methodologies, and practitioner writeups — each with its source. The pattern in all of them: the contract prices a number of GPUs, and the disputes are always about everything the number doesn't say.
Capacity guarantee vs uptime SLA conflation
Providers advertise 'Capacity guarantee: Yes' without percentages, measurement windows, or remedies. A capacity guarantee governs access to your reserved allocation; an uptime SLA governs whether it works. Demand exact percentages, per-node vs cluster-averaged measurement windows, and specific credits for misses — don't infer a number. [source]
Idle billing and hidden utilization floors
Reserved contracts bill flat hourly rates whether the cluster runs or sits idle; some contracts price the discount against a guaranteed minimum utilization, converting a capacity option into a usage obligation. Flex reservations (holding fee + usage rate) reduce idle penalty. Breakeven: a 45% discount only beats on-demand above ~55% utilization. [source]
Uncapped overage rates referencing floating market prices
'Standard on-demand rates' in an overage clause is a pointer to whatever the provider charges that day, not a number. Negotiate a multiplier ceiling (e.g., 1.5x reserved rate max) and a pre-agreed 10-20% burst allowance instead of open-ended exposure during shortage spikes. [source]
Egress fees and undefined exit mechanics
Hyperscalers impose high data-transfer fees that function as switching costs (most neoclouds do not charge egress). Negotiate capped or waived egress at term end, defined data export formats, and dated migration assistance — e.g., NVIDIA's 30-day transition window as a concrete, dated obligation rather than goodwill language. [source]
Capacity reclaim / preemption clauses on discounted compute
Some providers sell discounted idle compute but retain the right to reclaim it with seven days' notice if a higher-paying customer appears — fine for fault-tolerant batch work, disastrous mid-training-run if you didn't read the clause. [source]
Hardware substitution rights breaking ASC 842 lease treatment
Multi-tenant products often retain rights to swap underlying hardware, which can fail the identified-asset test and disqualify the contract from ASC 842 lease accounting — changing balance-sheet treatment. Contracts under 12 months typically avoid capitalization; finance teams can use accounting treatment as leverage against 3-year locks. [source]
Late delivery with no escape clause
Late GPU deliveries are widespread in the GPU cloud industry; MSAs must specify exact delivery dates with escape clauses (or penalties) for delays, since the customer's own downstream commitments start on the promised date. [source]
Prepayment funding the provider's own buildout
Contracts over one year typically require ~20% of total contract value prepaid, and neocloud suppliers use these prepayments to finance their own infrastructure deals — meaning buyers carry counterparty/completion risk on capacity that may not exist yet. [source]
No-cancellation policies and penalty asymmetry
AWS Capacity Blocks have a strict no-cancellation policy; Google allows cancellation only before provisioning; most neocloud contracts impose significant financial penalties. Secondary-market resale of unused reservations is an emerging escape valve. Also watch auto-renew on multi-year commitments without explicit sign-off. Negotiable levers: step-down clauses at 12/24-month checkpoints, hour banking, and SKU-generation (not specific SKU) commitments allowing H100->H200 swaps. [source]
SLA-compliant but unusable clusters (link-flap loophole)
NIC flaps lasting one microsecond every minute can cause NCCL stalls of minutes to half-hours, yet some providers claim SLA compliance because nodes technically stayed 'up.' SLAs must define health in terms of workload-relevant metrics (flap rates, fabric error thresholds), not node ping. [source]
SLA credit mechanics: future credits vs real-time deductions
Credit mechanisms vary — some providers issue future bill credits (locking you in to collect your remedy), others deduct downtime from the current bill in real time; ClusterMAX 2.0 flags real-time deductions and explicit breach penalties as what buyers should demand. [source]
Discount tiers vs term-length lock-in math
March 2026 benchmarks: 6-month terms run 20-30% below on-demand, 1-year 40-50%, 3-year 55-72%; providers resist sub-6-month terms on premium GPUs. Multi-vendor bidding cuts 15-20%, demand consolidation 15-30%, and quarter-end timing improves leverage — but the discount only pays if utilization exceeds (1 - discount%). [source]
How the tracked providers stack up
The same clauses, provider by provider, from each seller's own published terms (full detail + sources on every provider page). n/p = not published — the seller states it nowhere we could find. Absence is a finding: an unpublished SLA is the link-flap loophole with better manners.
| Provider | Uptime SLA | Remedy | Capacity guarantee | Egress | Self-serve |
|---|---|---|---|---|---|
| CoreWeave | >=99.9% monthly uptime for instances dep… | credits — tiered financial credits on the monthly bill. Comp… | Yes, published in capacity-plans docs: 'Reserved Instances (… | Free — pricing page lists 'Data transfer… | no |
| Lambda | n/p | none stated — Terms of Service contain no uptime SLA or serv… | No published guarantee language. Pricing page describes on-d… | Free — billing docs state 'you are not c… | yes |
| Crusoe | 99.5% Monthly Uptime Percentage for GPU-… | credits — 'Financial Credits' applied to future bills: 10% o… | Reservations = 'an active reserved instance agreement, typic… | Free — 'At this time, Crusoe Cloud does … | yes |
| Nebius | 99.50% Service Uptime for Compute Cloud … | credits — compensation as a discount against service fees in… | Not published as a guarantee. Reserved/committed capacity ('… | No general compute/VM egress charge publ… | yes |
| Voltage Park | 99.5% monthly uptime ('>= 99.5%') — same… | credits — tiered Financial Credits: 10% of monthly bill for … | n/p | Free — pricing page states 'No hidden in… | yes |
| Together AI | GPU Clusters: none published — Terms of … | none stated — no credit/refund/remedy terms published on pro… | PTU: 'Reserved capacity means your traffic doesn't compete..… | Free — docs: 'Together AI does not charg… | partial |
| FluidStack | 99% network availability per calendar mo… | credits — 'Service credits are calculated on a monthly basis… | n/p | n/p | no |
| Verda | n/p | none stated in published terms — Terms & Conditions prov… | n/p | partially published: container registry … | yes |
| AWS | 99.99% region-level; 99.5% instance-leve… | credits — 10% credit for uptime 99.0-99.99%, 30% for 95.0-98… | No published 'guarantee' wording. EC2 Capacity Blocks for ML… | 100 GB/month data transfer out to intern… | partial — credit-card signup yes, but pu… |
| Microsoft Azure | 99.99% (2+ VMs across 2+ Availability Zo… | credits — multi-instance: <99.99%/<99.95% = 10%, <9… | Explicitly published but NOT for flagship AI GPUs: On-demand… | First 100 GB/month free; internet egress… | partial — pay-as-you-go credit-card sign… |
| Google Cloud | >=99.99% multi-zone; >=99.95% single ins… | credits — Financial Credits on future bills, sole and exclus… | Published AI Hypercomputer consumption options: standard res… | Premium Tier internet egress from US: ~1… | partial — credit-card signup works, but … |
| Oracle Cloud/OCI | >=99.9% Monthly Uptime Service Commitmen… | credits — exclusive remedy, claim-based. Multi-AD: <99.99… | Capacity reservations published: 'Assurance that you have th… | First 10 TB/month outbound free (all lis… | partial — Pay As You Go signup is self-s… |
| Hyperstack | 99.5% minimum uptime — but the same SLA … | credits — pro-rata rebate to account for fees paid in the af… | Published Reservation Pricing on the GPU pricing page (per-m… | Free — pricing page: 'Egress/Ingress tra… | yes — credit card via Stripe ('add your … |
| RunPod | No public SLA for self-serve tiers; '99.… | none stated publicly — ToS: 'all Fees for the Service are no… | Enterprise/sales-led only: 'Reserved Clusters: Dedicated GPU… | Free — docs state Pods have 'no fees for… | partial — Pods/Serverless fully self-ser… |
| Vast.ai | n/p | none stated — ToS affirmatively disclaims availability: 'Com… | On-demand instances 'Cannot be interrupted' (host-set fixed … | Host-set per-host bandwidth pricing, sho… | yes — credit cards via Stripe, plus cryp… |
| Scaleway | 99.5% SLO for GPU Instances (tiers: Shar… | credits (voucher on future invoices, 'cannot be reimbursed i… | Published reservation offers: 'GPU Cluster On Demand' (reser… | GPU instances page claims no data-moveme… | yes — on-demand GPU instances (incl. H10… |
| OVHcloud | 99.99% monthly availability for Public C… | credits — 10% of monthly charge if availability between 95.0… | No published GPU capacity guarantee. Savings Plans (12 month… | Included/free in most regions — inbound … | yes — immediate signup with US$200 free … |
| TensorDock | 99.99% marketed as a host standard ('Ten… | credits (account deductions) per published Downtime Compensa… | None published. On-demand has fixed pricing; spot instances … | 'no ingress/egress fees' (homepage claim… | yes for on-demand (deposit funds, deploy… |
| Prime Intellect | n/p | none stated contractually; ad-hoc credits/refunds per docs —… | On-demand pods are non-interruptible vs Spot pods interrupti… | n/p | yes for on-demand (dashboard signup, Str… |
As published 2026-09-14 · truncated for scanning — the provider pages carry full text and source links.
The through-line
Nearly every gotcha above is an unspecified deliverable: uptime measured on nodes while the fabric flaps, capacity "guaranteed" without a number, hardware swapped under an identified-asset test, overage priced against a floating reference. This is exactly what a delivery specification exists to fix — ASSAY-1 defines the record a contract can cite so "8×H100" stops being a negotiation and starts being a specification.
And the number every clause above orbits — what deals like yours actually close at — is exactly what list prices can't tell you. The street-price pilot aggregates redacted real prints, contribute-to-see.
Negotiating a GPU contract?
The weekly report tracks real prices and what sellers actually publish — the leverage is free.
Subscribe free