macrostack
Browse

The AI stack

Categories

Local & Sovereign AINotes & KnowledgeObservability & MonitoringPassword ManagersWeb AnalyticsTeam ChatSmart HomeNetworking & RoutersVideo ConferencingCloud Storage & SyncPhotos & MediaAPI DevelopmentImage EditingWorkflow Automation & iPaaSDeveloper Tools & ContainersOffice & Productivity SuitesNo-Code DatabasesCode Hosting & Git ForgesProject ManagementEmail Marketing & NewslettersScheduling & BookingError Tracking & Exception MonitoringLog Management & SIEMVPN & PrivacyEmail & Secure MailVector Databases & AI SearchLLM & Agent FrameworksDomains & Web HostingData Removal & PrivacyAuthentication & IdentityHelp Desk & Customer SupportCloud & VPSKubernetes & Container PlatformsEmbedding ModelsPDF & DocumentsAI Coding AssistantsAI Voice & SpeechLLM Observability & EvaluationLLM Gateways & RoutingCloud GPU & AI ComputeCI/CD & build automationData & pipeline orchestrationModel serving & inferenceAI agent frameworksBackend as a serviceSecrets managementFeature flags & experimentationProduct analyticsSearch infrastructureUptime & status monitoringAffiliate & partner platformsVisitor identification & personalisationWikis & internal docsIdentity & access managementData warehouses & analytics enginesCustomer data platformsCRMObject storageBI & dashboardsE-signatureWhiteboards & diagrammingIn-memory data stores & cachingPlatform as a serviceTransactional & bulk emailHeadless CMSDesign & prototypingE-commerce platformsInternal tools & admin panelsManaged databasesForms & surveysFine-Tuning & Model TrainingRAG & Retrieval PlatformsLLM Evaluation & TestingAI Guardrails & Content SafetySpeech Recognition & TranscriptionExperiment Tracking & ML OpsDocument AI & OCRCompliance automation & security posture

About

How we rank & score

Layer 2.5 · head to head

AWS (P5, P6, G6) vs Google Cloud (A3, A4)

Capacity Blocks against Dynamic Workload Scheduler — two answers to the same shortage.

AWS (P5, P6, G6)Google Cloud (A3, A4)
Can you leaveMovable with effort. Expect to redo images, storage wiring and networking.Movable with effort. Expect to redo images, storage wiring and networking.
CounterpartyEffectively none. This is the reason the premium exists.Effectively none.
What it isA general cloud that also rents accelerators. Most expensive per hour, and the one your compliance team has already approved.A general cloud that also rents accelerators. Most expensive per hour, and the one your compliance team has already approved.
How you reach itYou get virtual machines. The familiar cloud model.You get virtual machines. The familiar cloud model.
AcceleratorsH100 (P5), B200 (P6), L4/L40S (G6), plus Trainium and InferentiaH100, H200, B200, plus TPU v5e/v6/v7
RegionsGlobalGlobal
Pricing modelOn-demand, spot, Savings Plans, and Capacity Blocks for reserved GPU windows.On-demand, spot, committed use discounts, Dynamic Workload Scheduler.
Getting startedExisting AWS account.Existing GCP account.
CapacityConstrained on the newest parts; Capacity Blocks exist precisely because on-demand cannot be relied on.Constrained on newest parts; DWS is the queueing mechanism.
OwnershipAmazon.Alphabet.

AWS (P5, P6, G6)

For: Anyone whose data, VPC, compliance boundary and team already live in AWS. The GPU price is rarely the deciding number.

The catch: Reported around $9.36/GPU-hour for a B200 Capacity Block against roughly $5.50 at Lambda. You are buying integration and counterparty certainty, and paying for both.

Economics: Egress and adjacency costs usually dominate the GPU line. Compare total workload cost, never the hourly rate alone.

Google Cloud (A3, A4)

For: Teams already on GCP, and anyone who wants TPUs — which exist nowhere else.

The catch: TPU is the real differentiator and the real lock-in. Choosing a TPU is choosing Google Cloud for the life of that workload; there is no second supplier.

Economics: Spot B200 reported around $6.69/GPU-hour. Committed-use discounts change the picture substantially and are where the actual negotiation happens.

Neither table row is a price

Deliberately. Published on-demand rates at this layer move weekly, and essentially nobody signing a real contract pays them — every serious buyer pays less than every list figure either of these companies publishes. Quoting one here would date this page within a month.

The GPU rental price index carries dated, sourced figures instead, and the durable finding there is the spread: the identical H100 rents from roughly $1.38 to $12.29 an hour depending only on who you rent it from.

The layers underneath both

Whichever you pick is renting you chips in a building that needs power. In 2026 that is the constraint that binds: Microsoft has disclosed an Azure backlog it cannot fill for want of megawatts rather than accelerators, and the US interconnection queue exceeds 2,600 GW with roughly 80% of projects withdrawing before they energise.

Layer 2 — Silicon · Layer 1 — Energy · The interconnection queue

Verified 2026-09-09. We do not benchmark clusters and take no position paid for by either company. Where a provider here runs a referral programme it has not moved its placement — the comparison was written before any link was attached.

The Macrostack brief

New swaps, worth your inbox.

A short, occasional email when we add a high-intent alternative or ship a new head-to-head. No spam, no selling your address — unsubscribe in one click.