Heterogeneous vGPU#

Heterogeneous vGPU allows a single physical GPU to simultaneously support multiple vGPU profiles with different memory allocations (framebuffer sizes). This configuration is beneficial for environments where VMs have diverse GPU resource requirements. By enabling the same physical GPU to host vGPUs of varying sizes, heterogeneous vGPU optimizes overall resource usage, ensuring VMs access only the necessary GPU resources and preventing underutilization.

When a GPU is configured for heterogeneous vGPU, its behavior during events like a host reboot, NVIDIA Virtual GPU Manager reload, or GPU reset varies by hypervisor.

Note

Heterogeneous vGPU configuration only supports the Best Effort and Equal Share schedulers.

Heterogeneous vGPU is supported on Volta and later GPUs. For additional information and operational instructions across different hypervisors, refer to the Heterogeneous vGPU documentation.

Platform Support for Heterogeneous vGPUs#

NVIDIA AI Enterprise Infra Release: NVIDIA AI Enterprise Infra 7.x

Refer to Configuring a GPU for Heterogeneous vGPU on RHEL KVM.

NVIDIA AI Enterprise Infra Release: NVIDIA AI Enterprise Infra 7.x

Refer to Configuring a GPU for Heterogeneous vGPU on Linux KVM.

NVIDIA AI Enterprise Infra Release: NVIDIA AI Enterprise Infra 7.x

Refer to Configuring a GPU for Heterogeneous vGPU on VMware vSphere.