Heterogeneous Devices
GPUStack uses one request model for GPUs, NPUs, MLUs, DCUs and PPUs. The resource name and the runtime isolation method depend on the manufacturer.
Contents
Accelerator requests
Accelerator Requests
explains whole, shared, logically sliced and
physically partitioned requests. Device Discovery
explains
how the per-node Devices record connects discovery to allocation.
Device operations
Check a node before installation with Preflight Operations . If you use hardware partitioning, follow the runbook for NVIDIA , T-Head or Hygon .