UC-01Use cases

Where modular AI infrastructure fits — and where it does not

Modular AI infrastructure fits organizations that need dedicated GPU capacity on their own terms — their site, their power, their data — and cannot wait years for a conventional facility. It fits poorly where demand is small or intermittent, where no realistic power path exists, or where an existing data center still has headroom.

8
Use-case profiles
4
Recurring disqualifiers
6
Checks before shortlisting
PODOS Pod sited behind a university research buildingCONCEPTUAL VISUALIZATION

The pattern behind every profile

01

What the workload actually is

Sustained fine-tuning and steady inference behave differently from bursty experimentation, and only one of them justifies owned capacity.

02

Which constraint binds first

Power, cooling, space, data governance, or time. The binding one decides whether hardware selection is even the right next step.

03

Relief, or relocation

Whether a factory-built unit relieves that constraint — or merely relocates it somewhere the problem is still yours.

Framing

Broad-based demand, meeting rooms built for another decade

Demand for dedicated AI capacity is broad-based, not a hyperscaler phenomenon. U.S. data centers consumed about 4.4% of national electricity in 2023, with Lawrence Berkeley National Laboratory projecting 6.7–12% by 2028[1], and the IEA expects global data-centre electricity use to roughly double toward ~945 TWh by 2030, driven largely by AI[2]. Yet the facilities most organizations already operate were not built for this: the Uptime Institute's 2025 operator survey places typical rack densities in the 10–30 kW band, well below what dense GPU nodes draw [3].

Every profile below therefore turns on the same three questions: what the workload actually is, which constraint binds first — power, cooling, space, data governance, or time — and whether a factory-built unit like the one described on the platform overview relieves that constraint or merely relocates it. These are workload profiles and design intent, not customer references: PODOS is an early-stage company, and no deployments, customers, or certifications are claimed on this page.

Vertical guides

Four verticals have a dedicated guide

Each guide takes one profile further: the workload in detail, the roles that must agree, the site requirements, and the cases where the answer is no.

Router

Fit at a glance

A strong fit signal is a reason to read the matching profile; a weak fit signal is a reason to stop before spending engineering time.

CodeProfileTypical workloadBinding constraintStrong fit signalWeak fit signal
U-01Enterprise AIFine-tuning + steady inferenceRack density in existing roomsHigh sustained GPU utilizationSpiky, experimental demand
U-02Universities & researchShared training queuesCampus power and coolingFunded multi-year cluster demandHPC hall with headroom
U-03ManufacturingInspection + digital twinsData gravity at plant sitesMV service on the estateOne-rack inference load
U-04HealthcareImaging + clinical language modelsData control, constrained estatesInstitution-controlled landCompliance path not yet scoped
U-05Government & secureSovereign or air-gapped AIAuthorization timelinesDefined-perimeter requirementData cleared for certified cloud
U-06EdgeRegional inference servingThin megawatt-class middle tierMetro power near demandKilowatt-scale sites
U-07Supplemental capacityGPU expansion of a full facilityLive-hall retrofit disruptionCampus land plus spare powerStranded in-hall capacity
U-08Power producersCompute at the generation sourceInterconnection queuesCurtailed or queued megawattsIntermittent supply, no storage

Profiles

Use-case profiles

Each profile states both sides of the fit question — where a factory-built unit helps, and where it does not.

U-01

Enterprise AI

WorkloadSustained fine-tuning and internal inference on proprietary data — the pattern where reserved cloud GPU commitments run at high utilization month after month.Binding constraintCorporate server rooms were provisioned for single-digit-kW racks[3], and retrofitting one for direct-to-chip liquid cooling is often a larger project than the cluster itself.Where modular helpsA dedicated unit on company or leased industrial land gives the cluster a purpose-built envelope. Each PODOS Pod is designed as a standardized 1 MW building block, designed for 128 GPUs, so capacity planning stays arithmetic — units, not bespoke halls.Where it does notBursty experimentation and workloads that idle most of the week. If a megawatt cannot be kept busy, shared cloud remains the honest default.

U-02

Universities & research

WorkloadMany principal investigators sharing one scheduler; long training queues; hardware funded in discrete grant awards.Binding constraintCampus machine rooms rarely hold spare megawatts or liquid-cooling loops, and new academic buildings move at capital-planning speed.Where modular helpsA self-contained unit sited near existing campus electrical infrastructure hands facilities teams a fixed, documented envelope — power in, heat out — instead of an open-ended construction program.Where it does notInstitutions whose HPC hall still has power and cooling headroom; expanding in place is usually cheaper. Funding rules that favor operating spend over capital purchases also point back to cloud credits.

U-05

Government & secure

WorkloadTraining and inference under sovereignty, air-gap, or controlled-access requirements that rule out shared cloud regions.Binding constraintProcurement and facility-authorization timelines dominate; a program can hold budget for compute yet wait years for an approved place to put it.Where modular helpsA physically bounded unit gives security teams a defined perimeter to assess — one enclosure, one power feed, documented ingress — rather than a shared hall.Where it does notModular construction shortens the building, not the authorization. PODOS claims no government accreditation, and programs whose data can lawfully run in certified cloud regions may find that path faster.

U-06

Edge

WorkloadRegional inference serving — models placed near users or data sources rather than in a distant region.Binding constraintThe middle tier is thin: device-level edge AI and hyperscale regions are both well served, while megawatt-class capacity in secondary metros is scarce.Where modular helpsA unit designed to be relocatable can occupy that middle tier where metro power exists — and move if demand does.Where it does notTrue edge sites measured in kilowatts — a closet rack, not a pod. Any latency benefit is workload-specific and should be measured before committing; no general number honestly applies. Full edge AI guide.

U-07

Supplemental capacity

WorkloadAn operator whose existing halls are out of power or cooling headroom for GPU racks, with demand still arriving.Binding constraintRetrofitting a live hall for high-density liquid cooling disrupts tenants and takes floor space out of service.Where modular helpsAdded capacity beside the existing facility, on the same campus and network, while the main hall keeps running. PODOS targets a 90-day window from order to commissioning for a standard unit — the deployment process is documented separately.Where it does notHalls with stranded power and empty white space are often better served by a targeted retrofit — run that comparison before adding enclosures.

U-08

Power producers & stranded energy

WorkloadTurning curtailed, queued, or under-contracted generation into a sellable compute product by colocating AI capacity at the source.Binding constraintInterconnection queues run years, and the IEA reports grid-connection bottlenecks tightening even as data-centre electricity use surged in 2025 [5].Where modular helpsCompute placed behind the meter consumes power where it is generated, and NREL has demonstrated data centers operating as flexible grid assets, including a 70 MW grid-interactive facility [4]. A unit-sized building block is intended to be matched to generation blocks rather than forcing a monolithic campus.Where it does notSites without a workable fiber path, or highly intermittent generation without storage. Compute economics depend on sustained utilization; a resource that runs a few hundred hours a year cannot carry a cluster.

U-03 · Manufacturing

Data gravity keeps the compute on the estate

Workload. Vision inspection, defect detection, digital-twin simulation, and process optimization — heavy inference near the line plus periodic retraining on plant telemetry.

Binding constraint. Data gravity. Plants generate more camera and sensor data than is economical to backhaul, and industrial estates have power but no data hall.

Where modular helps. Many industrial sites already take medium-voltage service — the class of input the pod's power architecture is designed around — so a unit can sit on the same estate as the machines it serves.

Where it does not. Batch analytics that tolerate a round trip to a cloud region, or a single line whose inference fits in one rack. A megawatt is the wrong granularity for a kilowatt problem.

PODOS Pod in the yard of a modern manufacturing plantCONCEPTUAL VISUALIZATION

U-04 · Healthcare

A utility pad, not clinical space

Workload. Imaging models, clinical documentation, and research on protected records — cases where governance teams want data on infrastructure the institution controls.

Binding constraint. Hospital estates are chronically short on space and power, and clinical buildings are the wrong place for a GPU cluster.

Where modular helps. A dedicated unit on institution-controlled property keeps training and inference inside the organization's own physical and network perimeter.

Where it does not. PODOS claims no healthcare compliance certification. Regulatory review, privacy assessment, and accreditation are the operator's work and run on their own clock; smaller inference loads may also fit hardware the institution already owns. The healthcare guide takes this further.

PODOS Pod on a utility pad beside a modern hospital buildingCONCEPTUAL VISUALIZATION

The question is never whether a factory-built unit is impressive. It is whether it relieves the binding constraint — or merely relocates it.

UC-01 · The test applied to all eight profiles

8

Profiles, both sides stated

HONEST LIMITS

Where modular does not fit

Four disqualifiers recur across every vertical, and they are worth naming plainly.

  • Small or spiky demand. A megawatt-class unit is the wrong tool for teams still in evaluation, and shared cloud absorbs that phase better.
  • The absence of a power path. A pod does not create electrons, so a site still needs medium-voltage service, behind-the-meter generation, or a credible interconnection position.
  • Existing headroom. An organization with usable space, power, and cooling in a facility it already runs should price a retrofit before pricing new enclosures — the modular vs traditional comparison treats this honestly.
  • Operations. Staffing, monitoring, and maintenance do not disappear because the building arrived on a truck.
  • PODOS is an early-stage company: capacity and timeline figures on this page are design targets, not measured results from operating deployments, and no customer installations, certifications, or accreditations are claimed.
  • Regulated buyers in particular should treat compliance as their own workstream with its own calendar.

Decision

Six checks before shortlisting modular

  1. 01Utilization. Can you keep a megawatt-class unit productive most of the year — or aggregate workloads across teams until you can?
  2. 02Power path. Do you have existing medium-voltage service, behind-the-meter generation, or a credible interconnection position?
  3. 03Site. Is there a pad, delivery access, and a fiber route your network team accepts?
  4. 04Data reason. Is there a governance, gravity, or sovereignty reason to leave shared cloud — or only a cost hypothesis?
  5. 05Operating model. Who runs the unit on day 2, and with what monitoring and maintenance arrangement?
  6. 06The alternative. Have you scored modular against a retrofit of what you have and against traditional construction?

If most checks pass, the next steps are the PODOS Pod unit page for what a unit is, the engineering section for how its systems work, and the AI infrastructure glossary for the vocabulary used across this site.

PODOS Pod at a remote edge site with solar array and telecom mastCONCEPTUAL VISUALIZATION

Sources

  1. [1] 2024 United States Data Center Energy Usage Report (LBNL-2001637)Lawrence Berkeley National Laboratory, Dec 2024
  2. [2] Energy and AI — Executive SummaryInternational Energy Agency, Apr 2025
  3. [3] Global Data Center Survey 2025Uptime Institute, Jul 2025
  4. [4] Demonstrating the Data Center as a Flexible Grid AssetNREL (U.S. Department of Energy), FY2025
  5. [5] Data centre electricity use surged in 2025 even with tightening bottlenecksInternational Energy Agency, 2025

Test your profile against a real configuration

Bring the utilization curve, the binding constraint, and the power path. The configurator walks the same variables an engineering review would.

Size your deploymentSee the deployment process