</>macrostackBrowse all
The map

The AI stack, in five layers

AI is a layered cake. Electricity at the bottom, chips above it, the infrastructure that rents those chips by the hour, the models that run on it, and — at the top — the software people actually open. Every bill you pay is one of these five layers charged back to you.

Almost every buyer's guide describes one slice of one layer. The few maps of the whole stack are published by companies that sell one of the layers. This one is drawn by nobody who sells any of them, and it is wired to 48 real decisions with 204 verified alternatives underneath.

L5 Applications

27 decisions · 110 alternatives

The software people actually open — the layer where AI stops being infrastructure and becomes a product.

The largest layer by headcount and the most crowded. It is also where subscription creep happens fastest, because every app is now $10 to $30 a seat a month and nobody audits the total.

Who dominates it: Microsoft 365 · Google Workspace · Notion · Adobe · Salesforce

L4 Models & tooling

7 decisions · 35 alternatives

The models themselves, plus the gateways, vector stores, frameworks and observability that make them usable.

The fastest-moving layer, and the one where open weights have most closed the gap on closed ones. A decision made here in 2024 is almost certainly wrong now.

Who dominates it: OpenAI · Anthropic · Google DeepMind · Meta Llama · Mistral · Qwen

L3 Infrastructure

14 decisions · 59 alternatives

The clouds, orchestration and plumbing that turn silicon into something you can rent by the hour.

This is where the money leaks. Nearly every recurring bill a technical team complains about — GPU hours, CI minutes, log ingestion, cluster management — is billed at this layer, and nearly every one of them has a free engine underneath the paid convenience.

Who dominates it: AWS · Google Cloud · Azure · RunPod · Vast.ai · GitHub

L2 Chips & silicon

not covered

The physical accelerators — GPUs, TPUs, custom inference silicon — and the fabs that make them.

The single scarcest input in the industry. Which chip you can actually get, at what price, decides what you can afford to run.

Who dominates it: NVIDIA · AMD · Google TPU · AWS Trainium · TSMC

Not covered yet. Buying silicon is a procurement decision made by a handful of companies, not a choice a reader makes — but consumer GPUs for local AI are a real decision, and that is the honest way in if we ever open this layer.

L1 Energy

not covered

The electricity and cooling that everything above it consumes.

A datacentre's real constraint in 2026 is not chips, it is megawatts and the grid connection to deliver them. Every price further up the stack is downstream of this number.

Who dominates it: Utility grids · Nuclear PPAs · Gas turbines · Solar + storage

We do not cover this layer, and we are unlikely to. There is no consumer or developer decision here — you cannot choose a power contract the way you choose a database. We show it because leaving it out would make the map wrong, not because we intend to fill it.

How to read this map

The substitution is different at every layer, and that is the whole point of drawing it. At the application layer you are usually replacing a per-seat subscription with software you host. At the infrastructure layer you are usually paying a convenience premium over an open engine that is already inside the product you rent — a managed Airflow is Airflow, a managed inference platform is very often vLLM. At the model layer the open weights have closed most of the gap and the decision is now about where you run them, not whether they are good enough.

Two layers are blank here and stay blank. We show them because leaving them out would make the map flattering rather than true, and because the shape of the whole cake explains prices further up it: a GPU hour costs what it costs because of megawatts and fab capacity, neither of which you can shop for.

Every figure on the pages below is verified against primary sources and dated. Rankings are merit-only — affiliate income never changes an order or a pick. See our methodology.

The Macrostack brief

New swaps, worth your inbox.

A short, occasional email when we add a high-intent alternative or ship a new head-to-head. No spam, no selling your address — unsubscribe in one click.