macrostack
Browse

The AI stack

Categories

Local & Sovereign AINotes & KnowledgeObservability & MonitoringPassword ManagersWeb AnalyticsTeam ChatSmart HomeNetworking & RoutersVideo ConferencingCloud Storage & SyncPhotos & MediaAPI DevelopmentImage EditingWorkflow Automation & iPaaSDeveloper Tools & ContainersOffice & Productivity SuitesNo-Code DatabasesCode Hosting & Git ForgesProject ManagementEmail Marketing & NewslettersScheduling & BookingError Tracking & Exception MonitoringLog Management & SIEMVPN & PrivacyEmail & Secure MailVector Databases & AI SearchLLM & Agent FrameworksDomains & Web HostingData Removal & PrivacyAuthentication & IdentityHelp Desk & Customer SupportCloud & VPSKubernetes & Container PlatformsEmbedding ModelsPDF & DocumentsAI Coding AssistantsAI Voice & SpeechLLM Observability & EvaluationLLM Gateways & RoutingCloud GPU & AI ComputeCI/CD & build automationData & pipeline orchestrationModel serving & inferenceAI agent frameworksBackend as a serviceSecrets managementFeature flags & experimentationProduct analyticsSearch infrastructureUptime & status monitoringAffiliate & partner platformsVisitor identification & personalisationWikis & internal docsIdentity & access managementData warehouses & analytics enginesCustomer data platformsCRMObject storageBI & dashboardsE-signatureWhiteboards & diagrammingIn-memory data stores & cachingPlatform as a serviceTransactional & bulk emailHeadless CMSDesign & prototypingE-commerce platformsInternal tools & admin panelsManaged databasesForms & surveysFine-Tuning & Model TrainingRAG & Retrieval PlatformsLLM Evaluation & TestingAI Guardrails & Content SafetySpeech Recognition & TranscriptionExperiment Tracking & ML OpsDocument AI & OCRCompliance automation & security posture

About

How we rank & score
The map

The AI stack, in five layers

AI is a layered cake. Electricity at the bottom, chips above it, the infrastructure that rents those chips by the hour, the models that run on it, and — at the top — the software people actually open. Every bill you pay is one of these five layers charged back to you.

Almost every buyer's guide describes one slice of one layer. The few maps of the whole stack are published by companies that sell one of the layers. This one is drawn by nobody who sells any of them, and it is wired to 84 real decisions with 382 verified alternatives underneath.

L5 Applications

42 decisions · 183 alternatives

The software people actually open — the layer where AI stops being infrastructure and becomes a product.

The largest layer by headcount and the most crowded. It is also where subscription creep happens fastest, because every app is now $10 to $30 a seat a month and nobody audits the total.

Who dominates it: Microsoft 365 · Google Workspace · Notion · Adobe · Salesforce

L4 Models & tooling

14 decisions · 70 alternatives

The models themselves, plus the gateways, vector stores, frameworks and observability that make them usable.

The fastest-moving layer, and the one where open weights have most closed the gap on closed ones. A decision made here in 2024 is almost certainly wrong now.

Who dominates it: OpenAI · Anthropic · Google DeepMind · Meta Llama · Mistral · Qwen

L3 Infrastructure

28 decisions · 129 alternatives

The clouds, orchestration and plumbing that turn silicon into something you can rent by the hour.

This is where the money leaks. Nearly every recurring bill a technical team complains about — GPU hours, CI minutes, log ingestion, cluster management — is billed at this layer, and nearly every one of them has a free engine underneath the paid convenience.

Who dominates it: AWS · Google Cloud · Azure · RunPod · Vast.ai · GitHub

L2.5 Compute

covered — dedicated guide

The companies that actually rent you the accelerators — neoclouds, GPU marketplaces, serverless inference and the hyperscalers.

Layer 2 tells you which chip. Layer 3 tells you which platform. Neither answers the question with the money attached: from whom do you rent it, and what happens to your workload if they are not there in two years. A multi-year GPU commitment is a credit decision wearing a cloud contract.

Who dominates it: CoreWeave · Lambda · Nebius · RunPod · Vast.ai · Together AI

Open the Computeguide →

L2 Chips & silicon

covered — dedicated guide

The physical accelerators — GPUs, TPUs, custom inference silicon — and the fabs that make them.

The single scarcest input in the industry. Which chip you can actually get, at what price, decides what you can afford to run.

Who dominates it: NVIDIA · AMD · Google TPU · AWS Trainium · TSMC

Open the Chips & siliconguide →

L1 Energy

covered — dedicated guide

The electricity and cooling that everything above it consumes.

A datacentre's real constraint in 2026 is not chips, it is megawatts and the grid connection to deliver them. Every price further up the stack is downstream of this number.

Who dominates it: Utility grids · Nuclear PPAs · Gas turbines · Solar + storage

Open the Energyguide →

How to read this map

The substitution is different at every layer, and that is the whole point of drawing it. At the application layer you are usually replacing a per-seat subscription with software you host. At the infrastructure layer you are usually paying a convenience premium over an open engine that is already inside the product you rent — a managed Airflow is Airflow, a managed inference platform is very often vLLM. At the model layer the open weights have closed most of the gap and the decision is now about where you run them, not whether they are good enough.

Two layers are blank here and stay blank. We show them because leaving them out would make the map flattering rather than true, and because the shape of the whole cake explains prices further up it: a GPU hour costs what it costs because of megawatts and fab capacity, neither of which you can shop for.

Every figure on the pages below is verified against primary sources and dated. Rankings are merit-only — affiliate income never changes an order or a pick. See our methodology.

The Macrostack brief

New swaps, worth your inbox.

A short, occasional email when we add a high-intent alternative or ship a new head-to-head. No spam, no selling your address — unsubscribe in one click.