ComputeAssay

Buyer's guide · compiled 2026-09-10

The GPU contract gotcha list

Reserved GPU contracts are where the deliverability problem becomes a legal problem. Below: the clauses that burn buyers, compiled from published negotiation guides, provider rating methodologies, and practitioner writeups — each with its source. The pattern in all of them: the contract prices a number of GPUs, and the disputes are always about everything the number doesn't say.

Capacity guarantee vs uptime SLA conflation

Providers advertise 'Capacity guarantee: Yes' without percentages, measurement windows, or remedies. A capacity guarantee governs access to your reserved allocation; an uptime SLA governs whether it works. Demand exact percentages, per-node vs cluster-averaged measurement windows, and specific credits for misses — don't infer a number. [source]

Idle billing and hidden utilization floors

Reserved contracts bill flat hourly rates whether the cluster runs or sits idle; some contracts price the discount against a guaranteed minimum utilization, converting a capacity option into a usage obligation. Flex reservations (holding fee + usage rate) reduce idle penalty. Breakeven: a 45% discount only beats on-demand above ~55% utilization. [source]

Uncapped overage rates referencing floating market prices

'Standard on-demand rates' in an overage clause is a pointer to whatever the provider charges that day, not a number. Negotiate a multiplier ceiling (e.g., 1.5x reserved rate max) and a pre-agreed 10-20% burst allowance instead of open-ended exposure during shortage spikes. [source]

Egress fees and undefined exit mechanics

Hyperscalers impose high data-transfer fees that function as switching costs (most neoclouds do not charge egress). Negotiate capped or waived egress at term end, defined data export formats, and dated migration assistance — e.g., NVIDIA's 30-day transition window as a concrete, dated obligation rather than goodwill language. [source]

Capacity reclaim / preemption clauses on discounted compute

Some providers sell discounted idle compute but retain the right to reclaim it with seven days' notice if a higher-paying customer appears — fine for fault-tolerant batch work, disastrous mid-training-run if you didn't read the clause. [source]

Hardware substitution rights breaking ASC 842 lease treatment

Multi-tenant products often retain rights to swap underlying hardware, which can fail the identified-asset test and disqualify the contract from ASC 842 lease accounting — changing balance-sheet treatment. Contracts under 12 months typically avoid capitalization; finance teams can use accounting treatment as leverage against 3-year locks. [source]

Late delivery with no escape clause

Late GPU deliveries are widespread in the GPU cloud industry; MSAs must specify exact delivery dates with escape clauses (or penalties) for delays, since the customer's own downstream commitments start on the promised date. [source]

Prepayment funding the provider's own buildout

Contracts over one year typically require ~20% of total contract value prepaid, and neocloud suppliers use these prepayments to finance their own infrastructure deals — meaning buyers carry counterparty/completion risk on capacity that may not exist yet. [source]

No-cancellation policies and penalty asymmetry

AWS Capacity Blocks have a strict no-cancellation policy; Google allows cancellation only before provisioning; most neocloud contracts impose significant financial penalties. Secondary-market resale of unused reservations is an emerging escape valve. Also watch auto-renew on multi-year commitments without explicit sign-off. Negotiable levers: step-down clauses at 12/24-month checkpoints, hour banking, and SKU-generation (not specific SKU) commitments allowing H100->H200 swaps. [source]

SLA-compliant but unusable clusters (link-flap loophole)

NIC flaps lasting one microsecond every minute can cause NCCL stalls of minutes to half-hours, yet some providers claim SLA compliance because nodes technically stayed 'up.' SLAs must define health in terms of workload-relevant metrics (flap rates, fabric error thresholds), not node ping. [source]

SLA credit mechanics: future credits vs real-time deductions

Credit mechanisms vary — some providers issue future bill credits (locking you in to collect your remedy), others deduct downtime from the current bill in real time; ClusterMAX 2.0 flags real-time deductions and explicit breach penalties as what buyers should demand. [source]

Discount tiers vs term-length lock-in math

March 2026 benchmarks: 6-month terms run 20-30% below on-demand, 1-year 40-50%, 3-year 55-72%; providers resist sub-6-month terms on premium GPUs. Multi-vendor bidding cuts 15-20%, demand consolidation 15-30%, and quarter-end timing improves leverage — but the discount only pays if utilization exceeds (1 - discount%). [source]

How the tracked providers stack up

The same clauses, provider by provider, from each seller's own published terms (full detail + sources on every provider page). n/p = not published — the seller states it nowhere we could find. Absence is a finding: an unpublished SLA is the link-flap loophole with better manners.

ProviderUptime SLARemedyCapacity guaranteeEgressSelf-serve
CoreWeave>=99.9% monthly uptime for instances dep…credits — tiered financial credits on the monthly bill. Comp…Yes, published in capacity-plans docs: 'Reserved Instances (…Free — pricing page lists 'Data transfer…no
Lambdan/pnone stated — Terms of Service contain no uptime SLA or serv…No published guarantee language. Pricing page describes on-d…Free — billing docs state 'you are not c…yes
Crusoe99.5% Monthly Uptime Percentage for GPU-…credits — 'Financial Credits' applied to future bills: 10% o…Reservations = 'an active reserved instance agreement, typic…Free — 'At this time, Crusoe Cloud does …yes
Nebius99.50% Service Uptime for Compute Cloud …credits — compensation as a discount against service fees in…Not published as a guarantee. Reserved/committed capacity ('…No general compute/VM egress charge publ…yes
Voltage Park99.5% monthly uptime ('>= 99.5%') — same…credits — tiered Financial Credits: 10% of monthly bill for …n/pFree — pricing page states 'No hidden in…yes
Together AIGPU Clusters: none published — Terms of …none stated — no credit/refund/remedy terms published on pro…PTU: 'Reserved capacity means your traffic doesn't compete..…Free — docs: 'Together AI does not charg…partial
FluidStack99% network availability per calendar mo…credits — 'Service credits are calculated on a monthly basis…n/pn/pno
Verdan/pnone stated in published terms — Terms & Conditions prov…n/ppartially published: container registry …yes
AWS99.99% region-level; 99.5% instance-leve…credits — 10% credit for uptime 99.0-99.99%, 30% for 95.0-98…No published 'guarantee' wording. EC2 Capacity Blocks for ML…100 GB/month data transfer out to intern…partial — credit-card signup yes, but pu…
Microsoft Azure99.99% (2+ VMs across 2+ Availability Zo…credits — multi-instance: <99.99%/<99.95% = 10%, <9…Explicitly published but NOT for flagship AI GPUs: On-demand…First 100 GB/month free; internet egress…partial — pay-as-you-go credit-card sign…
Google Cloud>=99.99% multi-zone; >=99.95% single ins…credits — Financial Credits on future bills, sole and exclus…Published AI Hypercomputer consumption options: standard res…Premium Tier internet egress from US: ~1…partial — credit-card signup works, but …
Oracle Cloud/OCI>=99.9% Monthly Uptime Service Commitmen…credits — exclusive remedy, claim-based. Multi-AD: <99.99…Capacity reservations published: 'Assurance that you have th…First 10 TB/month outbound free (all lis…partial — Pay As You Go signup is self-s…
Hyperstack99.5% minimum uptime — but the same SLA …credits — pro-rata rebate to account for fees paid in the af…Published Reservation Pricing on the GPU pricing page (per-m…Free — pricing page: 'Egress/Ingress tra…yes — credit card via Stripe ('add your …
RunPodNo public SLA for self-serve tiers; '99.…none stated publicly — ToS: 'all Fees for the Service are no…Enterprise/sales-led only: 'Reserved Clusters: Dedicated GPU…Free — docs state Pods have 'no fees for…partial — Pods/Serverless fully self-ser…
Vast.ain/pnone stated — ToS affirmatively disclaims availability: 'Com…On-demand instances 'Cannot be interrupted' (host-set fixed …Host-set per-host bandwidth pricing, sho…yes — credit cards via Stripe, plus cryp…
Scaleway99.5% SLO for GPU Instances (tiers: Shar…credits (voucher on future invoices, 'cannot be reimbursed i…Published reservation offers: 'GPU Cluster On Demand' (reser…GPU instances page claims no data-moveme…yes — on-demand GPU instances (incl. H10…
OVHcloud99.99% monthly availability for Public C…credits — 10% of monthly charge if availability between 95.0…No published GPU capacity guarantee. Savings Plans (12 month…Included/free in most regions — inbound …yes — immediate signup with US$200 free …
TensorDock99.99% marketed as a host standard ('Ten…credits (account deductions) per published Downtime Compensa…None published. On-demand has fixed pricing; spot instances …'no ingress/egress fees' (homepage claim…yes for on-demand (deposit funds, deploy…
Prime Intellectn/pnone stated contractually; ad-hoc credits/refunds per docs —…On-demand pods are non-interruptible vs Spot pods interrupti…n/pyes for on-demand (dashboard signup, Str…

As published 2026-09-14 · truncated for scanning — the provider pages carry full text and source links.

The through-line

Nearly every gotcha above is an unspecified deliverable: uptime measured on nodes while the fabric flaps, capacity "guaranteed" without a number, hardware swapped under an identified-asset test, overage priced against a floating reference. This is exactly what a delivery specification exists to fix — ASSAY-1 defines the record a contract can cite so "8×H100" stops being a negotiation and starts being a specification.

And the number every clause above orbits — what deals like yours actually close at — is exactly what list prices can't tell you. The street-price pilot aggregates redacted real prints, contribute-to-see.

Negotiating a GPU contract?

The weekly report tracks real prices and what sellers actually publish — the leverage is free.

Subscribe free