macrostack
Browse

The AI stack

Categories

Local & Sovereign AINotes & KnowledgeObservability & MonitoringPassword ManagersWeb AnalyticsTeam ChatSmart HomeNetworking & RoutersVideo ConferencingCloud Storage & SyncPhotos & MediaAPI DevelopmentImage EditingWorkflow Automation & iPaaSDeveloper Tools & ContainersOffice & Productivity SuitesNo-Code DatabasesCode Hosting & Git ForgesProject ManagementEmail Marketing & NewslettersScheduling & BookingError Tracking & Exception MonitoringLog Management & SIEMVPN & PrivacyEmail & Secure MailVector Databases & AI SearchLLM & Agent FrameworksDomains & Web HostingData Removal & PrivacyAuthentication & IdentityHelp Desk & Customer SupportCloud & VPSKubernetes & Container PlatformsEmbedding ModelsPDF & DocumentsAI Coding AssistantsAI Voice & SpeechLLM Observability & EvaluationLLM Gateways & RoutingCloud GPU & AI ComputeCI/CD & build automationData & pipeline orchestrationModel serving & inferenceAI agent frameworksBackend as a serviceSecrets managementFeature flags & experimentationProduct analyticsSearch infrastructureUptime & status monitoringAffiliate & partner platformsVisitor identification & personalisationWikis & internal docsIdentity & access managementData warehouses & analytics enginesCustomer data platformsCRMObject storageBI & dashboardsE-signatureWhiteboards & diagrammingIn-memory data stores & cachingPlatform as a serviceTransactional & bulk emailHeadless CMSDesign & prototypingE-commerce platformsInternal tools & admin panelsManaged databasesForms & surveysFine-Tuning & Model TrainingRAG & Retrieval PlatformsLLM Evaluation & TestingAI Guardrails & Content SafetySpeech Recognition & TranscriptionExperiment Tracking & ML OpsDocument AI & OCRCompliance automation & security posture

About

How we rank & score

Layer 2 · head to head

Google TPU v6e (Trillium) vs Google TPU v7 (Ironwood)

One generation apart, and Google publishes far more detail about the newer one — which is itself a signal about how they want them compared.

Google TPU v6e (Trillium)Google TPU v7 (Ironwood)
Can you get itCannot be bought. Exists only inside one cloud, so choosing it is choosing that cloud.Cannot be bought. Exists only inside one cloud, so choosing it is choosing that cloud.
MemoryNot published in directly comparable terms192 GB HBM3E per chip
BandwidthNot published7.37 TB/s per chip
ComputeNot published in comparable FP8 terms4,614 FP8 TFLOPS per chip
PowerNot publishedNot published
InterconnectICI9.6 Tb/s inter-chip (ICI)
SoftwareJAX, XLA, PyTorch/XLAJAX, XLA, PyTorch/XLA
Workloadbothinference

Google TPU v6e (Trillium)

For: The previous TPU generation, still widely provisioned on Google Cloud and cheaper per hour than Ironwood.

The catch: Same lock-in as any TPU: Google Cloud only. Google publishes far less per-chip detail for v6e than for v7, which is itself a signal about how they want it compared.

Economics: Priced as a cloud SKU. Compare against v7 on tokens delivered per dollar, not on any spec sheet.

Google TPU v7 (Ironwood)

For: Inference at scale on Google Cloud. Ships in 256-chip and 9,216-chip configurations.

The catch: You cannot buy this. It exists only inside Google Cloud, so choosing it is choosing a cloud, permanently. It is also inference-optimised — do not benchmark it as a training part.

Economics: Vertically integrated: the price you see is a cloud price, not a chip price, and it is not comparable to a $/GPU-hour rental line.

Before either — can you power it?

Not published against Not published. In 2026 that comparison usually matters more than the FLOPS one: the US interconnection queue exceeds 2,600 GW with waits approaching five years, and roughly 80% of projects withdraw before energising.

Layer 1 — Energy · Nuclear vs gas · Direct-to-chip cooling

Specifications verified 2026-09-06. We take no commission at this layer, on either part.

The Macrostack brief

New swaps, worth your inbox.

A short, occasional email when we add a high-intent alternative or ship a new head-to-head. No spam, no selling your address — unsubscribe in one click.