In this article:
Want us to find IT vendors for you?
Share your vendor requirements with one of our account managers, then we build a vetted shortlist and arrange introductory calls with each vendor.
Book a call

CoreWeave vs Lambda vs Nebius: renting US GPU capacity for enterprise AI

A vendor-neutral comparison of how CoreWeave, Lambda and Nebius sell reserved GPU capacity in the US, which GPUs run in which states, what federal and regulated buyers can use today, and when AWS, Azure or Google Cloud fits better.

Author:
Date

Price per GPU-hour leads almost every comparison of CoreWeave, Lambda and Nebius. It also tells you the least, because a rate on a pricing page says nothing about whether the GPUs you need will exist when your project starts.

What decides an enterprise rental is how each provider sells capacity, the smallest unit you can rent, which GPUs run where, and the terms behind all three. I compare those four things for US buyers, using each provider's own documentation and SEC filings, and I leave list prices out on purpose: they changed during my research, and none of them says whether capacity exists.

The US is where most of this capacity sits. CoreWeave lists 49 of its 59 availability zones in the US, Lambda runs nine of its 14 regions here, and Nebius has announced new US sites in New Jersey, Missouri, Alabama and Pennsylvania.

CoreWeave, Lambda and Nebius are GPU clouds, often called neoclouds: providers built around dedicated GPU capacity, with a narrower set of services than AWS, Azure or Google Cloud. This guide covers renting that capacity for training, fine-tuning and inference.

CoreWeave and Nebius also sell hosted model APIs, which sit outside this comparison. If you meant calling a hosted model, start with our AWS Bedrock vs Azure OpenAI vs Google Vertex AI comparison.

IT Leaders Report 2026

What are your IT peers investing in 2026?

We spoke to about 1,300 IT leaders from various organisations to understand what they're evaluating. Most of it is the kind of thing you'd only hear from a peer you know well enough to ask, so we've put it in one place.

Read the report
IT Leaders Report 2026 cover artwork

‍

On-demand access is not a capacity guarantee

On-demand GPU access is a place in a queue, and none of the three providers promises more than that. CoreWeave opens its General Access zones to every customer, subject to capacity, and keeps Dedicated Access zones for select customers.

Lambda describes its on-demand instances as self-serve, first-come access. Nebius applies default regional quotas to GPU virtual machines and InfiniBand use, and sends large cluster requests to its sales team.

The rest of the market says the same in writing. Microsoft's capacity reservation documentation states that quota and capacity are separate checks, and Crusoe's support documentation says quota approval does not guarantee capacity.

Four checks stand between a GPU request and running hardware

On-demand access clears the first three. Only a reservation adds the fourth.

1

Account approved

The provider accepts you as a customer.

Set by signup checks or sales

2

Quota granted

Permission to deploy a set number of GPUs in a region.

Set by the provider's quota limits

3

Capacity free

The hardware exists in that region when you ask for it.

Set by what other customers already hold

4

Capacity contracted

The hardware is held for you for a fixed term.

Set by your order form

On-demand covers checks 1 to 3. Check 3 is best effort at CoreWeave, Lambda and Nebius.

A reservation adds check 4. Its guarantee is only as strong as the terms you sign.

Sources: CoreWeave regions documentation, Lambda instances page, Nebius quota documentation, Microsoft Learn capacity reservation overview.

‍

In the US right now, that queue is long. On CoreWeave's Q1 2026 earnings call, its CFO said the company is largely sold out of 2026 capacity and has started allocating 2027.

Nebius reported selling out of available capacity in its Q3 2025 shareholder letter, and again in each of the next two quarters. On its Q1 2026 call, management said four or more customers typically compete for every GPU it brings online.

The largest buyers lock in capacity first. Lambda signed a multi-year agreement with Microsoft to run tens of thousands of NVIDIA GPUs, including GB300 NVL72 systems, in its US data centers.

US GPU capacity is spoken for before it comes online

CoreWeave

2026

Capacity largely sold out, with 2027 now being allocated

CFO, Q1 2026 earnings call, May 7, 2026

CoreWeave

~5 yrs

Weighted-average length of its committed, take-or-pay customer contracts

Form 10-K for FY2025

Nebius

3 qtrs

Sold out of available capacity in Q3 2025, Q4 2025 and Q1 2026

Shareholder letter and earnings calls

Nebius

4+

Customers typically competing for each GPU it brings online

Management, Q1 2026 earnings call

Statements cover each provider's whole fleet. Sources: CoreWeave Q1 2026 earnings call transcript and Form 10-K for FY2025; Nebius Q3 2025 shareholder letter and Q4 2025 and Q1 2026 earnings calls.

‍

Enterprise workloads end up on reservations for a plain reason. A training run that needs 128 GPUs for four months needs every one of them in the same region, on the same cluster network, for the whole term, and a first-come queue cannot promise that.

The first test I run with any GPU provider is one request: a written capacity commitment for our GPU type, GPU count, US region and start date. The answer, and how long it takes to arrive, tells me more than any pricing page. For the facility side of the same problem, see our guide to winning the power and capacity race.

Several claims repeated across current comparison pages fail a check against provider documentation.

ClaimWhat the documentation showsStatus
CoreWeave sells only reserved contractsReservations, Flex Reservations, On-Demand and Spot are all documented, and Spot became generally available in March 2026False
CoreWeave offers no single-GPU instancesThe GH200 rents as one GPU in Virginia and Nevada, and single-GPU inference rates apply to inference platform customersPartly true
CoreWeave is FedRAMP authorizedCoreWeave Federal, launched October 28, 2025, is pursuing FedRAMP and other authorizationsFalse
Nebius is a Europe-only cloudPublic US regions in Kansas City, Missouri and Woodbury, MinnesotaFalse
Nebius offers H100 clusters in the USIts US regions carry H200, B200 and B300; H100 runs in FinlandFalse
Lambda rents B200 only as 8-GPU serversB200 and H100 both come in 1, 2, 4 and 8 GPU sizes on demandFalse
Lambda has filed to go publicOnly Form D exempt-offering notices on SEC EDGAR, with no registration statementFalse

Checked against CoreWeave, Lambda and Nebius documentation, company announcements and SEC EDGAR on October 6, 2026.

‍

How CoreWeave, Lambda and Nebius are built

All three now run GPU clusters with managed Kubernetes and Slurm, and all three sell managed AI services on top. The differences that change a buying decision sit lower down: what you can rent, how you buy it and where in the US it runs.

CoreWeave runs bare-metal GPU servers through its own Kubernetes service and SUNK, its Slurm-on-Kubernetes product. Its H100, H200, B200 and B300 instances come as whole 8-GPU servers, and its rack-scale GB200 and GB300 NVL72 systems rent as 4-GPU instances. CoreWeave Forge adds managed inference, fine-tuning and Weights & Biases tooling, and CoreWeave trades on Nasdaq.

Lambda sells self-serve virtual machines with one, two, four or eight H100 or B200 GPUs, plus 1-Click Clusters of 16 to more than 2,000 GPUs. Superclusters and private cloud go through sales. Lambda offers managed Kubernetes and Slurm, has no spot tier, and is privately held.

Nebius rents virtual machines with one or eight GPUs on dedicated hosts, with managed Kubernetes and Soperator, its managed Slurm service. It adds managed PostgreSQL, managed MLflow and Token Factory for hosted inference. Nebius is headquartered in Amsterdam, trades on Nasdaq, and runs two of its nine public regions in the US.

The rentable unit matters more than it looks. A job that needs four GPUs in one machine gets a 4-GPU instance on Lambda and a full 8-GPU server on CoreWeave or Nebius.

The smallest unit each provider rents

CoreWeave

8 GPUs
4 (NVL72)
1 (GH200)

H100, H200, B200 and B300 rent as whole 8-GPU servers. GB200 and GB300 NVL72 rent as 4-GPU instances, and the GH200 rents as one GPU.

Lambda

1
2
4
8 GPUs
Clusters from 16

H100 and B200 rent in 1, 2, 4 or 8 GPUs on demand. 1-Click Clusters start at 16 GPUs.

Nebius

1 GPU
8 GPUs

1 or 8 GPUs per VM. Clusters are built from 8-GPU hosts.

Sources: CoreWeave GPU instance documentation, Lambda On-Demand Cloud documentation, Nebius Compute product page.

AttributeCoreWeaveLambdaNebius
Delivery modelBare-metal GPU serversVirtual machines and dedicated clustersVirtual machines on dedicated hosts
Smallest unit you can rent8-GPU server (GH200: 1 GPU)1 GPU1 GPU
Cluster productInfiniBand clusters on CKS or SUNK1-Click Clusters, 16 to 2,000+ GPUsGPU clusters on InfiniBand fabrics
OrchestrationCKS (Kubernetes) and SUNK (Slurm)Managed Kubernetes and managed SlurmManaged Kubernetes and Soperator (Slurm)
Software and managed servicesCoreWeave Forge: inference, fine-tuning, Weights & BiasesLambda Stack on every instance, managed Kubernetes and SlurmToken Factory, managed PostgreSQL, managed MLflow
Spot or preemptibleSpotNonePreemptible
Large capacity bought throughAccount team, on multi-year committed contractsSales, for terms over 1 year and SuperclustersSales team, up to five-year dedicated agreements
US footprint49 zones in 20 states, 18 open to all customers9 regions in VA, DC, IL, TX, CA, AZ and UTKansas City, MO and Woodbury, MN
Company statusPublic (Nasdaq: CRWV)PrivatePublic (Nasdaq: NBIS)

Sources: CoreWeave instance and zone documentation and Form 10-K; Lambda On-Demand Cloud documentation and pricing page; Nebius Compute and region documentation and Form 6-K. Checked October 6, 2026.

‍

If your team is still weighing Kubernetes against Slurm for these clusters, our EKS vs AKS vs GKE guide covers the managed Kubernetes trade-offs. Our WEKA vs VAST Data vs DDN vs Pure Storage comparison covers the storage that feeds a training run.

‍

How each provider sells reserved capacity

Lambda is the only one of the three that publishes its reservation terms. Its 1-Click Clusters run from 16 to more than 2,000 B200 or H100 GPUs on terms of two weeks to one year, and longer terms go to its sales team.

Even that published path has a human checkpoint. Lambda's billing documentation invoices a reservation once Lambda approves it, and on October 6, 2026, every cluster row on its pricing page routed to a sales conversation.

CoreWeave documents four capacity plans: Reservations, Flex Reservations, On-Demand and Spot. Its Form 10-K for 2025 says customers generally buy a specified amount of capacity on multi-year, take-or-pay contracts, which averaged about five years at the end of 2025.

Flex Reservations, in preview since March 10, 2026, guarantee capacity up to a ceiling you choose without committing you to run at that ceiling around the clock.

Nebius sells on-demand and preemptible capacity and multi-month cluster reservations, and its largest contracts are five-year dedicated agreements. Its Form 424B5 describes two dedicated GPU clusters for Meta over five years, alongside the Microsoft agreement below.

How each provider sells GPU capacity, from self-serve to dedicated agreements

Each chip is a documented way to buy. The further right it sits, the longer the term and the larger the commitment.

Longer terms, larger commitments

Self-serve

Reserved

Dedicated agreements

CoreWeave

On-DemandSpot
ReservationsFlex Reservations (preview)Multi-year take-or-pay contracts
31 single-tenant Dedicated Access zones in the US

Lambda

Instances, 1 to 8 GPUs
1-Click Clusters, 2 weeks to 1 yearCluster terms over 1 year
Superclusters and private cloudMulti-year Microsoft GB300 deployment

Nebius

On-demand and preemptible, within quotas
Multi-month cluster reservations
Five-year Microsoft capacity in Vineland, NJFive-year Meta clusters

Sources: CoreWeave capacity plans, zone list and Form 10-K; Lambda pricing page and Microsoft announcement; Nebius pricing page and SEC filings.

‍

The clearest picture of a large reservation comes from a US securities filing. Nebius's Form 6-K on its Microsoft agreement describes dedicated GPU capacity delivered in tranches from Vineland, New Jersey over five years, with service level commitments and liquidated damages for late delivery.

If Nebius misses a delivery date after a grace period and cannot provide alternative capacity, Microsoft can terminate that tranche. That is a capacity commitment, a remedy and a substitution clause in one paragraph, which makes it a useful template for your own order form.

Inside a documented US reservation contract

Nebius and Microsoft: dedicated GPU capacity from Vineland, New Jersey, as described in Nebius's September 8, 2025 Form 6-K

Five-year termCapacity delivered in tranchesService level commitments
Delivery date
Each tranche of GPU capacity has an agreed delivery date.
Grace period
A missed date starts a grace period before remedies apply.
Alternative capacity? Nebius can offer substitute capacity for a late tranche.

If it can
The termination right is not triggered.

If it cannot
Microsoft has the right to terminate that tranche.

Late delivery also carries liquidated damages, a pre-agreed payment for the delay.

Map it to your own order form

Capacity commitmentA delivery date for each block of GPUs you reserve
RemedyLiquidated damages for late delivery, plus a right to exit the late block
SubstitutionWhen the provider may offer alternative capacity, and whether you must accept it

Source: Nebius Group Form 6-K, September 8, 2025. Nebius's later Form 424B5 states the parties' obligations began in early November 2025.

‍

Where the GPUs sit in the US

Where a GPU runs decides latency to your data, which zones you can reach without a sales call, and how close your team is to the hardware it depends on. The three providers spread across the US very differently.

CoreWeave lists 49 US availability zones across 20 states. Eighteen zones in ten states are General Access, open to every customer, and the other 31 are Dedicated Access zones that CoreWeave describes as single-tenant, with no other customer sharing the zone or its network fabric.

Lambda runs nine US regions: Virginia, Washington DC, Illinois, three in Texas, California, Arizona and Utah. Lambda notes that instance types vary by region, so confirm the GPU and the region together.

Nebius runs two public US regions. Kansas City, Missouri carries B200 and H200 plus every managed service Nebius sells, while Woodbury, Minnesota carries B300 with a narrower set that leaves out managed Slurm, PostgreSQL and MLflow.

Where the three providers run GPU capacity in the US

CoreWeave: 49 zones in 20 statesLambda: 9 regionsNebius: 2 regions
CoreWeave zone open to all customersCoreWeave Dedicated Access onlyLambda regionNebius region
AK
ME
VT
NH
WA
ID
MT
ND
MN
IL
WI
MI
NY
RI
MA
OR
NV
WY
SD
IA
IN
OH
PA
NJ
CT
CA
UT
CO
NE
MO
KY
WV
VA
MD
DE
AZ
NM
KS
AR
TN
NC
SC
DC
OK
LA
MS
AL
GA
HI
TX
FL

Tile map, one square per state. Sources: CoreWeave All Availability Zones, Lambda On-Demand Cloud regions, Nebius AI Cloud regions. Checked October 6, 2026.

‍

GPU type narrows the map further. CoreWeave's availability matrix puts H100 in ten open-access US zones, while its newest rack-scale systems sit in fewer places: GB300 NVL72 in three zones and Vera Rubin NVL72 in one, in Texas.

Which GPUs run where in the US

CoreWeave

Open-access US zones by GPU, with states

H100

10
TXNYNJVAOHNVAZWA

H200

5
NYNJVAOHAZ

GB200 NVL72

5
NYNJOHMINV

B200

4
ILMIVAWA

B300

4
TXMINV

GB300 NVL72

3
ILTXNV

Vera Rubin NVL72

1
TX

Nebius and Lambda

US regions

Nebius, Kansas City, MO

us-central1

B200H200RTX PRO 6000

Managed Kubernetes, Soperator (Slurm), PostgreSQL, MLflow and Serverless AI

Nebius, Woodbury, MN

us-north1

B300

Managed Kubernetes, object storage and Container Registry

Lambda, nine US regions

VA, DC, IL, TX (3), CA, AZ, UT

B200, 1 to 8H100, 1 to 8GH200A100

Instance types vary by region

Sources: CoreWeave instance availability matrix and zone list; Nebius AI Cloud regions; Lambda On-Demand Cloud documentation. Checked October 6, 2026.

‍

Both GPU clouds are adding US capacity fast. Nebius is building a 300 MW data center in Vineland, New Jersey and has announced sites in Independence, Missouri and Birmingham, Alabama, plus a 1.2 GW campus in Pennsylvania. Lambda operates from 15 US data centers and is set to open a facility in Kansas City.

For US-only data, Lambda's cloud terms matter more than its region list.

Lambda's defaults, until your order form changes them

Data location

You consent to hosting in the US, and Lambda may move your data to its other regions, which include Japan, India, Israel and Germany.

Health data

Platform guidelines bar personal health information and other sensitive personal data.

Backups

Lambda has no obligation to maintain or back up your data.

Source: Lambda Cloud Terms of Service and Platform Guidelines, last updated August 2025.

‍

CoreWeave groups all of its US locations in a single US Geo and says it places regions to help meet data residency requirements. Nebius's data processing agreement commits to processing personal data in the region the customer chooses.

‍

Federal and regulated US workloads

Federal agencies and contractors handling controlled unclassified information need FedRAMP-authorized capacity, and today that runs through the hyperscalers. AWS brought Capacity Blocks for ML to GovCloud on June 12, 2026, with B200 in GovCloud (US-West) and B200 and B300 in GovCloud (US-East).

Microsoft added H200 GPUs to its Azure Secret and Top Secret clouds in December 2025. Both options run inside government-authorized boundaries.

CoreWeave is building the GPU-cloud path. It launched CoreWeave Federal on October 28, 2025 to pursue FedRAMP and other authorizations, and on July 30, 2026 it partnered with Leidos to deliver AI cloud services inside classified facilities.

Lambda's government page describes work with agencies through Cooperative Research and Development Agreements, backed by a US-based support team.

Federal GPU capacity, October 2025 to July 2026

GPU cloudHyperscaler

October 28, 2025CoreWeave

Launches CoreWeave Federal to pursue FedRAMP and other authorizations for agencies and the Defense Industrial Base.

December 4, 2025Microsoft Azure

Adds NVIDIA H200 GPUs to the Azure Government Secret and Top Secret clouds.

June 12, 2026AWS

Brings Capacity Blocks for ML to GovCloud: B200 in US-West, B200 and B300 in US-East, reservable for up to six months.

July 30, 2026CoreWeave

Announces a partnership with Leidos to deliver AI cloud services inside classified facilities.

Lambda works with agencies through Cooperative Research and Development Agreements, with a US-based support team.

Sources: CoreWeave and AWS announcements, Microsoft Azure Government blog, Lambda government page, Unite.AI report on the Leidos partnership.

‍

Regulated commercial data splits the field differently. Nebius offers a HIPAA business associate agreement on request and requires it before protected health information is uploaded, while Lambda's platform guidelines bar personal health information unless an order permits it.

For California consumer data, Lambda's trust page lists CCPA compliance, and Nebius's data processing agreement casts Nebius as a CCPA service provider.

‍

Enterprise readiness: SLAs, compliance and counterparty risk

An SLA tells you what a provider owes you when something breaks, and the three publish very different answers. CoreWeave's terms of service set a service level objective of 99.9% monthly uptime for instances in multiple regions, with service credits as the remedy.

Nebius publishes a 99.5% compute SLA for a single VM and makes compensation the sole remedy. Lambda's cloud terms provide the service as-is, with no uptime commitment.

An objective is a target. Read CoreWeave's credit terms before treating it as a commitment, and ask Lambda for an SLA in the order form.

All three hold SOC 2 Type II reports and ISO 27001 certification, shared through their trust portals (CoreWeave, Lambda, Nebius).

AttributeCoreWeaveLambdaNebius
Published uptime commitment99.9% objectiveInstances in multiple regionsNoneService provided as-is99.5% SLASingle VM
Remedy when it is missedService creditsNone publishedCompensation, as the sole remedy
SOC 2 Type IIYesYesYes
ISO 27001YesYesYes
HIPAA business associate agreementNot on Trust CenterBarred by defaultOn request
Federal pathCoreWeave FederalPursuing FedRAMPCRADA engagementsCommercial regions
SEC reportingFiles with the SEC (Nasdaq: CRWV)Private, Form D notices onlyFiles with the SEC (Nasdaq: NBIS)

Sources: CoreWeave terms of service, Trust Center and federal announcement; Lambda Cloud Terms of Service, Trust Portal and government page; Nebius SLA, Trust Center and HIPAA guideline; SEC EDGAR. Checked October 6, 2026.

‍

Counterparty risk comes down to what you can see and what the contract says. CoreWeave and Nebius file periodic reports with the SEC, which is why their contract structures appear in this guide. Lambda is private, and SEC EDGAR shows only exempt-offering notices for it, with no registration statement, as of October 6, 2026.

Lambda's terms also let it delete your data after termination without obligation and assign the contract in a merger without your consent. Whichever provider you choose, ask for data return windows, transition assistance and change-of-control rights in the order form.

‍

Why GPU capacity plans fail

GPU projects rarely stall on the GPU itself. They stall on assumptions made before anyone signs.

  • Planning on on-demand. First-come capacity is the first to disappear when providers report selling out, as CoreWeave and Nebius both did in 2026.
  • Choosing the provider before the GPU and the state. Nebius has no US H100, and CoreWeave's GB300 NVL72 runs in three US zones.
  • Signing the default terms. Default data-location and data-type clauses apply until your order form replaces them.
  • Renting a whole server for half a need. A four-GPU job on an 8-GPU unit leaves half the server idle.
  • Leaving hardware failure out of the contract. Large clusters lose nodes regularly, and the contract should say what happens next.

That last point deserves numbers. In The Llama 3 Herd of Models, Meta reported 419 unexpected interruptions during a 54-day pre-training snapshot on 16,384 H100 GPUs, and traced 58.7% of them to GPU issues.

What interrupted Meta's Llama 3 pre-training run

419unexpected interruptions

54 dayspre-training snapshot

16,384H100 GPUs in the cluster

90%+effective training time kept

Share of unexpected interruptions, by cause

Faulty GPU30.1%
GPU HBM3 memory17.2%
Network switch and cable8.4%
GPU SRAM memory4.5%
GPU system processor4.1%
GPU-related causes, 58.7% of unexpected interruptions in totalNetwork

Selected root-cause categories, as reported in Table 5 of Meta's The Llama 3 Herd of Models (2024). Bars are scaled to the largest category.

‍

Meta still kept more than 90% effective training time, with heavy automation behind it. Ask each provider how it detects and replaces a failed node, how quickly, and whether replacement time counts against the SLA.

For the budget side of the same planning, see why AI/ML workloads break cloud budgets.

‍

CoreWeave vs Lambda

CoreWeave and Lambda both sell InfiniBand clusters with managed Kubernetes and Slurm across many US locations, so the split comes down to how you buy and what you can see. Lambda publishes its cluster terms and rents H100 and B200 from a single GPU up.

CoreWeave sells whole servers on multi-year committed contracts and adds Spot and Flex Reservations for uneven demand. Its US map is wider and documented GPU by GPU, with 18 open zones in ten states, while Lambda's nine US regions cover six states and DC.

Verdict

Choose CoreWeave when you need rack-scale GB200 or GB300, a multi-year contracted cluster, or capacity in a specific US state. Choose Lambda when you want published cluster terms and one-to-eight-GPU access without a long sales cycle.

Choose CoreWeave when

  • You need GB200, GB300 or Vera Rubin rack-scale systems
  • You plan a multi-year committed contract
  • You want spot capacity or Flex Reservations

Choose Lambda when

  • You want to read cluster terms before talking to sales
  • You need one, two or four H100 or B200 GPUs
  • Your workload fits a two-week to one-year cluster

‍

CoreWeave vs Nebius

CoreWeave and Nebius are both public and both quote large reservations through sales, but their US footprints are far apart: 49 CoreWeave zones against two Nebius regions. Nebius's Kansas City region pairs B200 and H200 with managed PostgreSQL, MLflow and Slurm, and Woodbury adds B300.

Nebius rents single GPUs, publishes a compute SLA and offers a HIPAA business associate agreement. CoreWeave sells whole servers, runs H100 in ten open US zones and is building a federal path through CoreWeave Federal.

Verdict

Choose CoreWeave for H100, rack-scale systems or a federal roadmap across many US states. Choose Nebius for B200, H200 or B300 next to managed data services, health data under a BAA, or single-GPU VMs.

Choose CoreWeave when

  • You need H100 clusters in the US
  • You need a specific state or several US zones
  • You expect to need FedRAMP through CoreWeave Federal

Choose Nebius when

  • You want B200 or H200 in Kansas City next to managed PostgreSQL and MLflow
  • Health data under HIPAA is in scope
  • You want single-GPU VMs or preemptible capacity

‍

Lambda vs Nebius

Lambda covers more of the US, with nine regions against Nebius's two, while Nebius documents exactly which GPU runs in each region. Lambda publishes reservation terms from two weeks and rents H100; Nebius quotes reservations, sells preemptible capacity and publishes a compute SLA.

On regulated data they part ways. Nebius signs a BAA on request, while Lambda's default terms bar health data and let it move data between regions.

Verdict

Choose Lambda when you want H100 or B200 near your team in the US, with reservation terms you can read first. Choose Nebius when you need B300, a published SLA, preemptible capacity or health data under a BAA.

Choose Lambda when

  • You need H100 in the US
  • You want published cluster terms from two weeks to one year
  • You want a region close to your team, from California to Virginia

Choose Nebius when

  • You need B300 in the US
  • You process health data under HIPAA
  • You want preemptible capacity or a published SLA

‍

When AWS, Azure or Google Cloud is the better choice

A GPU cloud adds a second vendor, a second network boundary and a second set of terms. For some workloads, that overhead outweighs everything above.

A hyperscaler usually wins in five cases: work that needs FedRAMP today; inference next to data already in AWS, Azure or Google Cloud; committed spend you already need to draw down; procurement rules that require an existing vendor; and workloads that lean on many adjacent managed services. Retrieval-heavy assistants are the common example, and our vector database comparison covers that architecture.

Hyperscaler reservations are shorter and self-serve. AWS Capacity Blocks for ML reserve GPU instances up to eight weeks ahead for up to six months, in blocks of 1 to 64 instances that can be shared across accounts. Google Cloud's calendar-mode reservations hold GPU VMs for up to 90 days once Google approves the request.

Azure needs a closer read. Its capacity reservations guarantee capacity but don't cover the ND-series or NCads H100 v5 VMs that carry its AI GPUs, and Azure Reserved Instances carry no capacity guarantee.

How long each documented GPU reservation runs

Hyperscaler blocks run for days or months. The GPU clouds' largest contracts run for years.

Self-serve reservation windowMulti-year contract

AWSCapacity Blocks for ML, including GovCloud

1 day to 6 months, booked up to eight weeks ahead, 1 to 64 instances per block.

Google CloudCalendar-mode reservations

Up to 90 days, held once Google approves the request.

Lambda1-Click Clusters

2 weeks to 1 year published. Longer terms go through sales.

CoreWeaveCommitted contracts

Average length of about five years

Multi-year, take-or-pay contracts, per its Form 10-K for FY2025.

NebiusMicrosoft agreement

Five-year dedicated capacity agreement

Delivered in tranches from Vineland, New Jersey.

01 year2 years3 years4 years5 years

Linear scale to five years. Sources: AWS Capacity Blocks for ML documentation and GovCloud announcement; Google Cloud calendar-mode reservation documentation; Lambda pricing page; CoreWeave Form 10-K for FY2025; Nebius Form 6-K, September 8, 2025.

‍

GPU clouds still win where the documents show it: multi-year capacity on contract, published cluster terms from two weeks on Lambda, a 49-zone US footprint on CoreWeave, and managed Slurm on all three. If your data already lives in a hyperscaler, ask any GPU cloud about private connectivity before you compare anything else.

‍

Why Crusoe, RunPod and Together AI are out of scope

Three providers appear in most current coverage of this market and sit outside this comparison. Crusoe sells its high-density GPUs, including H100 SXM, B200 and GB200, only through reserved instance agreements, which makes it a cluster-first peer of CoreWeave and a candidate for a follow-on comparison.

RunPod is developer-first, with self-serve Instant Clusters of 16 to 64 GPUs and larger clusters through sales. Together AI leads with hosted inference, and its GPU clusters take reservations of 1 to 90 days.

‍

Which GPU cloud fits your workload?

Answer six questions about your workload. The tool marks each provider as a fit, ruled out, or dependent on terms you'll need in writing, using the US availability and terms above, and flags when a hyperscaler deserves a look first.

Which GPU cloud fits your workload?

Six questions. Each provider comes back as a fit, ruled out, or dependent on terms you will need in writing.

‍

Choosing on technical fit

For US buyers, the decision comes down to four facts: the GPU you need, the state it must run in, how long you need it, and what your data is allowed to do.

  • Choose CoreWeave when you need H100 across many US states, rack-scale GB200 or GB300, a multi-year contracted cluster, or a federal roadmap through CoreWeave Federal.
  • Choose Lambda when you want published cluster terms from two weeks, one to eight H100 or B200 GPUs on demand, and your data excludes health records.
  • Choose Nebius when you need B200 or H200 in Kansas City with managed data services, B300 in Minnesota, a HIPAA business associate agreement, or preemptible capacity.
  • Choose AWS, Azure or Google Cloud when the work needs FedRAMP today, or inference sits next to data you already hold there.
  • Pin the hosting region in your order form on Lambda for any US-only data requirement.

Before you sign with any of the three, get these terms in writing.

Get these three things in writing

1

A capacity commitment for your exact GPU type, count, US region and start date

Ask for a delivery date per block of GPUs and the remedy if one slips, the way Nebius's Microsoft agreement sets them out.

2

Data location and handling terms

Which US region holds your data, metadata, logs and backups, and which data types the terms allow.

3

Termination, data return and substitution rights

Whether the provider can offer alternative capacity, how long you have to retrieve data, and what happens on a change of control.

‍

If a provider won't put these three things in writing, you have your answer.

Shortlisting GPU cloud capacity?

Which provider fits depends on your GPU type, your US region and the terms you can get in writing. Shortlist pre-vetted AI infrastructure vendors on TechnologyMatch, matched to all three. You stay anonymous until you choose to talk, you pick who to talk to, and it's free for buyers.

Find GPU cloud vendors

FAQ

What is a neocloud?

A neocloud is a cloud provider built around dedicated GPU capacity for AI training and inference, with a narrower set of services than AWS, Azure or Google Cloud. CoreWeave, Lambda and Nebius are all neoclouds.

Can you rent a single GPU on CoreWeave?

Only the GH200, which CoreWeave runs in Virginia and Nevada. Its H100, H200, B200 and B300 instances rent as whole 8-GPU servers. Lambda rents H100 and B200 in 1, 2, 4 or 8 GPUs, and Nebius rents single-GPU VMs.

Where are CoreWeave's US data centers?

CoreWeave lists 49 US availability zones across 20 states. Its 18 General Access zones, open to all customers, are in Texas, Illinois, New York, New Jersey, Virginia, Ohio, Michigan, Nevada, Arizona and Washington. The other 31 are Dedicated Access zones reserved for select customers, in states including Georgia, Pennsylvania, Oregon and North Dakota.

Does Nebius have data centers in the US?

Yes. Nebius runs public regions in Kansas City, Missouri, with B200 and H200, and Woodbury, Minnesota, with B300. It is also building a 300 MW data center in Vineland, New Jersey, and has announced sites in Missouri, Alabama and Pennsylvania. Its H100 capacity runs in Finland.

Is Lambda suitable for enterprise workloads?

Lambda holds SOC 2 Type II and ISO 27001, runs nine US regions and publishes cluster terms from two weeks to one year. Its default terms offer no uptime commitment, bar health data and let it move data between regions, so set the SLA, data location and allowed data types in your order form.

Are CoreWeave, Lambda or Nebius FedRAMP authorized?

None of the three lists a FedRAMP authorization as of October 6, 2026. CoreWeave launched CoreWeave Federal in October 2025 to pursue FedRAMP and other authorizations, and Lambda works with agencies through Cooperative Research and Development Agreements. For FedRAMP High work today, AWS GovCloud reserves B200 and B300 through Capacity Blocks for ML.

How do reserved GPU contracts work?

A reservation holds specific hardware for you for a fixed term. CoreWeave's customers generally sign multi-year, take-or-pay contracts averaging about five years, and Lambda publishes cluster terms from two weeks to one year. Nebius's Microsoft agreement shows the larger shape: delivery dates per tranche, a grace period, and liquidated damages for late delivery.

Can you run HIPAA workloads on CoreWeave, Lambda or Nebius?

Nebius offers a HIPAA business associate agreement on request and requires it before protected health information is uploaded. Lambda's default terms bar personal health information unless an order permits it, and CoreWeave's Trust Center lists no HIPAA attestation, so ask its account team directly.