GPUStack Operator

Heterogeneous Devices

GPUStack uses one request model for GPUs, NPUs, MLUs, DCUs and PPUs. The resource name and the runtime isolation method depend on the manufacturer.

Contents

Accelerator requests

Accelerator Requests explains whole, shared, logically sliced and physically partitioned requests. Device Discovery explains how the per-node Devices record connects discovery to allocation.

Device operations

Check a node before installation with Preflight Operations . If you use hardware partitioning, follow the runbook for NVIDIA , T-Head or Hygon .