# Cluster management & GPU orchestration (out of scope for a parts site; listed for completeness)

Part ID: 11.02 · Layer 11 (Software, tools & services touching the hardware (optional))

Software that schedules AI workloads onto GPU clusters and manages the compute fleet, sitting outside the physical parts this site otherwise tracks.

Cluster management and orchestration software allocates jobs across GPU clusters, handling scheduling, fault tolerance, and utilization — the layer between the physical compute (Layer 1) and the AI workloads running on it. It is listed here for completeness because it is part of how the hardware in every other layer gets used, but it is a software category rather than a physical part, so it sits outside the scope of what this site otherwise catalogs.

**AI delta:** Larger, more expensive GPU clusters raise the cost of scheduling and utilization mistakes, increasing demand for orchestration software built specifically for AI training and inference.

## Companies

| Company | Role | Confidence | Ownership | Ticker | Note |
|---|---|---|---|---|---|
| Dell | service | inferred | public | DELL |  |
| HPE | service | inferred | public | HPE |  |
| Nvidia | service | inferred | public | NVDA | Base Command, Mission Control, Run:ai |
| Supermicro | service | inferred | public | SMCI |  |
| Rafay | service | inferred | private |  |  |
| SchedMD/Slurm | service | inferred | private |  |  |

Confidence tags: disclosed = company/filing statement; reported = trade press or partner list; inferred = taxonomy lead.
