macrostack
Browse

The AI stack

Categories

Local & Sovereign AINotes & KnowledgeObservability & MonitoringPassword ManagersWeb AnalyticsTeam ChatSmart HomeNetworking & RoutersVideo ConferencingCloud Storage & SyncPhotos & MediaAPI DevelopmentImage EditingWorkflow Automation & iPaaSDeveloper Tools & ContainersOffice & Productivity SuitesNo-Code DatabasesCode Hosting & Git ForgesProject ManagementEmail Marketing & NewslettersScheduling & BookingError Tracking & Exception MonitoringLog Management & SIEMVPN & PrivacyEmail & Secure MailVector Databases & AI SearchLLM & Agent FrameworksDomains & Web HostingData Removal & PrivacyAuthentication & IdentityHelp Desk & Customer SupportCloud & VPSKubernetes & Container PlatformsEmbedding ModelsPDF & DocumentsAI Coding AssistantsAI Voice & SpeechLLM Observability & EvaluationLLM Gateways & RoutingCloud GPU & AI ComputeCI/CD & build automationData & pipeline orchestrationModel serving & inferenceAI agent frameworksBackend as a serviceSecrets managementFeature flags & experimentationProduct analyticsSearch infrastructureUptime & status monitoringAffiliate & partner platformsVisitor identification & personalisationWikis & internal docsIdentity & access managementData warehouses & analytics enginesCustomer data platformsCRMObject storageBI & dashboardsE-signatureWhiteboards & diagrammingIn-memory data stores & cachingPlatform as a serviceTransactional & bulk emailHeadless CMSDesign & prototypingE-commerce platformsInternal tools & admin panelsManaged databasesForms & surveysFine-Tuning & Model TrainingRAG & Retrieval PlatformsLLM Evaluation & TestingAI Guardrails & Content SafetySpeech Recognition & TranscriptionExperiment Tracking & ML OpsDocument AI & OCRCompliance automation & security posture

About

How we rank & score

Layer 2.5 · head to head

RunPod vs Vast.ai

Curated marketplace against open bidding: the price gap is the reliability gap.

RunPodVast.ai
Can you leaveMovable with effort. Expect to redo images, storage wiring and networking.Movable with effort. Expect to redo images, storage wiring and networking.
CounterpartyLow exposure for you: you are renting by the second, so the switching cost of a supplier failing is hours, not quarters. That is the honest advantage of the marketplace model.You are renting from individuals and small operators. Assume any single machine disappears without notice and design for it.
What it isBrokers capacity somebody else owns. Cheapest headline rates, and the machine you get is not the machine you chose.Brokers capacity somebody else owns. Cheapest headline rates, and the machine you get is not the machine you chose.
How you reach itYou push a container image and it runs. No cluster to operate.You push a container image and it runs. No cluster to operate.
AcceleratorsH100, H200, A100, L40S, RTX 4090, RTX 5090, and a long consumer tailWhatever the market is offering — RTX 5090, 4090, PRO 6000, V100, H100 and a long tail
RegionsGlobal, community and secure cloudsGlobal, wherever a host has hardware
Pricing modelPer-second billing, on-demand and spot. Two tiers: 'Secure Cloud' in real datacentres, 'Community Cloud' on other people's hardware.Open bidding marketplace, interruptible and on-demand. The clearing price is public without an account.
Getting startedA few dollars.Cents.
CapacityGenerally good on consumer parts, variable on datacentre parts.Deep on consumer silicon, thin and volatile on datacentre parts.
OwnershipPrivate.Private.

RunPod

For: Fine-tuning, batch jobs, inference experiments, anyone whose workload can checkpoint and move.

The catch: Community Cloud is somebody's machine somewhere. For anything with a compliance story attached, that distinction is the whole decision, and it is easy to miss in the pricing table.

Economics: Reported around $2/hr for an H100 on-demand — roughly a third of CoreWeave list. Per-second billing genuinely matters for bursty work.

Vast.ai

For: Price-sensitive batch work, research, and anyone benchmarking what compute *should* cost before signing a contract.

The catch: Reliability is the price of the price. Hosts vary enormously, verification tiers matter, and a cheap unverified box that dies at hour nine is not cheap.

Economics: Consistently the market floor — reported around $1.49/hr for an H100 where hyperscalers ask $7. Its public API is the closest thing this market has to a spot index, which is why this site measures it.

Neither table row is a price

Deliberately. Published on-demand rates at this layer move weekly, and essentially nobody signing a real contract pays them — every serious buyer pays less than every list figure either of these companies publishes. Quoting one here would date this page within a month.

The GPU rental price index carries dated, sourced figures instead, and the durable finding there is the spread: the identical H100 rents from roughly $1.38 to $12.29 an hour depending only on who you rent it from.

The layers underneath both

Whichever you pick is renting you chips in a building that needs power. In 2026 that is the constraint that binds: Microsoft has disclosed an Azure backlog it cannot fill for want of megawatts rather than accelerators, and the US interconnection queue exceeds 2,600 GW with roughly 80% of projects withdrawing before they energise.

Layer 2 — Silicon · Layer 1 — Energy · The interconnection queue

Verified 2026-09-09. We do not benchmark clusters and take no position paid for by either company. Where a provider here runs a referral programme it has not moved its placement — the comparison was written before any link was attached.

The Macrostack brief

New swaps, worth your inbox.

A short, occasional email when we add a high-intent alternative or ship a new head-to-head. No spam, no selling your address — unsubscribe in one click.