macrostack
Browse

The AI stack

Categories

Local & Sovereign AINotes & KnowledgeObservability & MonitoringPassword ManagersWeb AnalyticsTeam ChatSmart HomeNetworking & RoutersVideo ConferencingCloud Storage & SyncPhotos & MediaAPI DevelopmentImage EditingWorkflow Automation & iPaaSDeveloper Tools & ContainersOffice & Productivity SuitesNo-Code DatabasesCode Hosting & Git ForgesProject ManagementEmail Marketing & NewslettersScheduling & BookingError Tracking & Exception MonitoringLog Management & SIEMVPN & PrivacyEmail & Secure MailVector Databases & AI SearchLLM & Agent FrameworksDomains & Web HostingData Removal & PrivacyAuthentication & IdentityHelp Desk & Customer SupportCloud & VPSKubernetes & Container PlatformsEmbedding ModelsPDF & DocumentsAI Coding AssistantsAI Voice & SpeechLLM Observability & EvaluationLLM Gateways & RoutingCloud GPU & AI ComputeCI/CD & build automationData & pipeline orchestrationModel serving & inferenceAI agent frameworksBackend as a serviceSecrets managementFeature flags & experimentationProduct analyticsSearch infrastructureUptime & status monitoringAffiliate & partner platformsVisitor identification & personalisationWikis & internal docsIdentity & access managementData warehouses & analytics enginesCustomer data platformsCRMObject storageBI & dashboardsE-signatureWhiteboards & diagrammingIn-memory data stores & cachingPlatform as a serviceTransactional & bulk emailHeadless CMSDesign & prototypingE-commerce platformsInternal tools & admin panelsManaged databasesForms & surveysFine-Tuning & Model TrainingRAG & Retrieval PlatformsLLM Evaluation & TestingAI Guardrails & Content SafetySpeech Recognition & TranscriptionExperiment Tracking & ML OpsDocument AI & OCRCompliance automation & security posture

About

How we rank & score

Layer 1 · The numbers that decide

Tokens per watt

The metric that actually matters in 2026: useful model output per unit of power. It joins the chip layer to the energy layer, which is why both are on this site.

The number

What it is
The metric that actually matters in 2026: useful model output per unit of power. It joins the chip layer to the energy layer, which is why both are on this site.
Typical values
Vendor-published only, and never on a common benchmark

Verified 2026-09-06.

The catch

There is no neutral, standardised tokens-per-watt benchmark. Every published figure comes from a company selling one side of it. Treat all of them as marketing until an independent body publishes a method.

What this is powering

A GB200 NVL72 rack draws 120–132 kW and a GB300 pushes 135–200 kW, against a 2026 average rack of about 27 kW. The silicon decision and the power decision are the same decision, made eighteen months apart.

Layer 2 — Silicon · GB200 NVL72 · The 150 W inference option

Others in the numbers that decide

Sources

Macrostack sells nothing at this layer and takes no commission on any option here. That is precisely why this comparison did not exist — everyone qualified to write it sells one of the answers.

The Macrostack brief

New swaps, worth your inbox.

A short, occasional email when we add a high-intent alternative or ship a new head-to-head. No spam, no selling your address — unsubscribe in one click.