Layer 2.5 · head to head
Lambda vs RunPod
Real datacentre hourly against per-second marketplace billing.
| Lambda | RunPod | |
|---|---|---|
| Can you leave | Movable with effort. Expect to redo images, storage wiring and networking. | Movable with effort. Expect to redo images, storage wiring and networking. |
| Counterparty | Private, so you are diligencing a company that does not have to tell you anything. Reported among the group carrying GPU-collateralised debt maturing 2026-2028. | Low exposure for you: you are renting by the second, so the switching cost of a supplier failing is hours, not quarters. That is the honest advantage of the marketplace model. |
| What it is | Purpose-built GPU cloud. Owns or leases its own datacentre capacity and sells it directly. | Brokers capacity somebody else owns. Cheapest headline rates, and the machine you get is not the machine you chose. |
| How you reach it | You get virtual machines. The familiar cloud model. | You push a container image and it runs. No cluster to operate. |
| Accelerators | H100, H200, B200, GH200, A100 | H100, H200, A100, L40S, RTX 4090, RTX 5090, and a long consumer tail |
| Regions | US | Global, community and secure clouds |
| Pricing model | On-demand by the hour, plus reserved clusters. Pricing is published on the website, which is rarer at this layer than it should be. | Per-second billing, on-demand and spot. Two tiers: 'Secure Cloud' in real datacentres, 'Community Cloud' on other people's hardware. |
| Getting started | Credit card, single GPU, minutes. | A few dollars. |
| Capacity | Frequently sold out on the newest parts. Availability is the constraint, not price. | Generally good on consumer parts, variable on datacentre parts. |
| Ownership | Private, widely reported as IPO-track. | Private. |
Lambda
For: Researchers and small teams who want a real GPU in the next ten minutes without a procurement conversation.
The catch: The thing you want is often unavailable. Reserved capacity solves it and turns the credit-card product into a contract.
Economics: Among the lowest published rates on the newest silicon — reported lowest on B200 in an August 2026 comparison. Cheap when you can get it.
RunPod
For: Fine-tuning, batch jobs, inference experiments, anyone whose workload can checkpoint and move.
The catch: Community Cloud is somebody's machine somewhere. For anything with a compliance story attached, that distinction is the whole decision, and it is easy to miss in the pricing table.
Economics: Reported around $2/hr for an H100 on-demand — roughly a third of CoreWeave list. Per-second billing genuinely matters for bursty work.
Neither table row is a price
Deliberately. Published on-demand rates at this layer move weekly, and essentially nobody signing a real contract pays them — every serious buyer pays less than every list figure either of these companies publishes. Quoting one here would date this page within a month.
The GPU rental price index carries dated, sourced figures instead, and the durable finding there is the spread: the identical H100 rents from roughly $1.38 to $12.29 an hour depending only on who you rent it from.
The layers underneath both
Whichever you pick is renting you chips in a building that needs power. In 2026 that is the constraint that binds: Microsoft has disclosed an Azure backlog it cannot fill for want of megawatts rather than accelerators, and the US interconnection queue exceeds 2,600 GW with roughly 80% of projects withdrawing before they energise.
Layer 2 — Silicon · Layer 1 — Energy · The interconnection queue