macrostack
Browse

The AI stack

Categories

Local & Sovereign AINotes & KnowledgeObservability & MonitoringPassword ManagersWeb AnalyticsTeam ChatSmart HomeNetworking & RoutersVideo ConferencingCloud Storage & SyncPhotos & MediaAPI DevelopmentImage EditingWorkflow Automation & iPaaSDeveloper Tools & ContainersOffice & Productivity SuitesNo-Code DatabasesCode Hosting & Git ForgesProject ManagementEmail Marketing & NewslettersScheduling & BookingError Tracking & Exception MonitoringLog Management & SIEMVPN & PrivacyEmail & Secure MailVector Databases & AI SearchLLM & Agent FrameworksDomains & Web HostingData Removal & PrivacyAuthentication & IdentityHelp Desk & Customer SupportCloud & VPSKubernetes & Container PlatformsEmbedding ModelsPDF & DocumentsAI Coding AssistantsAI Voice & SpeechLLM Observability & EvaluationLLM Gateways & RoutingCloud GPU & AI ComputeCI/CD & build automationData & pipeline orchestrationModel serving & inferenceAI agent frameworksBackend as a serviceSecrets managementFeature flags & experimentationProduct analyticsSearch infrastructureUptime & status monitoringAffiliate & partner platformsVisitor identification & personalisationWikis & internal docsIdentity & access managementData warehouses & analytics enginesCustomer data platformsCRMObject storageBI & dashboardsE-signatureWhiteboards & diagrammingIn-memory data stores & cachingPlatform as a serviceTransactional & bulk emailHeadless CMSDesign & prototypingE-commerce platformsInternal tools & admin panelsManaged databasesForms & surveysFine-Tuning & Model TrainingRAG & Retrieval PlatformsLLM Evaluation & TestingAI Guardrails & Content SafetySpeech Recognition & TranscriptionExperiment Tracking & ML OpsDocument AI & OCR

About

How we rank & score

Layer 4 · self-hosting reality check

What it actually takes to self-host Docling

The docs say 2 GB. In practice you want 8 GB — the layout models are the memory cost. Here is the honest version — real requirements, real monthly cost, what you will be maintaining, and the one thing that catches people out.

Usually reached from AWS Textract alternatives, where Docling is one of the picks.

RAM — documented minimum2 GB
RAM — what it really needs8 GB — the layout models are the memory cost
CPU4 vCPU; a GPU speeds large batches considerably
Disk2 GB for models plus your documents
Monthly cost$0 on your own hardware. A 100,000-page corpus is a weekend of compute against a four-figure Textract invoice.
Setup time1 hour
How you install itpip install docling; models download on first run
Ongoing maintenanceLow.
Where it stops scalingParallelises across CPU cores well. Throughput is bounded by page complexity, not volume.

The thing that catches people out

The first run silently downloads gigabytes of layout and table models, so a container built without warming that cache re-downloads on every cold start — which turns a 3-second job into a 3-minute one in production. Bake the models into your image or mount a persistent cache directory.

When not to self-host Docling

Your corpus is handwritten. Textract is meaningfully better there and Docling is not close — route handwriting separately rather than migrating it.

Every guide here carries this section. A site that only ever tells you to self-host is selling something — the useful answer is sometimes no.

Other Layer 4 self-hosting guides

Common questions

How much RAM does Docling actually need?
8 GB — the layout models are the memory cost in practice. The documented minimum is 2 GB, which is the figure at which the process starts rather than the figure at which it works under real use. 4 vCPU; a GPU speeds large batches considerably alongside it.
What does self-hosting Docling cost per month?
$0 on your own hardware. A 100,000-page corpus is a weekend of compute against a four-figure Textract invoice. This is commodity VPS pricing and excludes your time, which is the larger cost for most people — budget for low.
How long does it take to set up Docling?
1 hour, via pip install docling; models download on first run.
When should I NOT self-host Docling?
Your corpus is handwritten. Textract is meaningfully better there and Docling is not close — route handwriting separately rather than migrating it.
What is the most common mistake when self-hosting Docling?
The first run silently downloads gigabytes of layout and table models, so a container built without warming that cache re-downloads on every cold start — which turns a 3-second job into a 3-minute one in production. Bake the models into your image or mount a persistent cache directory.
The Macrostack brief

New swaps, worth your inbox.

A short, occasional email when we add a high-intent alternative or ship a new head-to-head. No spam, no selling your address — unsubscribe in one click.