country_code
Skip to main content
Ctrl+K
NVIDIA AI Enterprise - Home NVIDIA AI Enterprise - Home

NVIDIA AI Enterprise

  • Documentation Home
NVIDIA AI Enterprise - Home NVIDIA AI Enterprise - Home

NVIDIA AI Enterprise

  • Documentation Home

Table of Contents

  • Quick Start Guide
  • Support Matrix
    • 8.2 (Latest)
    • 8.1
    • 8.0

Releases

  • Release Notes
    • 8.2 (Latest)
    • 8.1
    • 8.0
  • Upgrading from 8.1 to 8.2

NVIDIA vGPU for Compute

  • Overview
  • Limitations
  • Features
    • MIG-Backed vGPU
    • Device Groups
    • GPUDirect RDMA and GPUDirect Storage
    • Heterogeneous vGPU
    • Live Migration
    • Multi-vGPU and P2P
      • Multi-vGPU
      • Multi-vGPU Board Support
      • Peer-To-Peer (P2P) CUDA Transfers
    • NVIDIA NVSwitch
    • NVLink Multicast
    • Scheduling Policies
    • Suspend-Resume
    • Unified Virtual Memory (UVM)
      • UVM Board Support
    • Choosing a vGPU Mode for Inference
  • Installation
    • Host Setup
    • Installing NVIDIA vGPU Guest Driver
    • AI Enterprise Workload Setup
  • Licensing
  • Configuration
    • Host MIG Provisioning
    • Guest MIG Reconfiguration
  • NVIDIA vGPU Types by Hardware
    • Blackwell Architecture vGPU Types
    • Hopper Architecture vGPU Types
      • Hopper H800 vGPU Types
      • Hopper H200 vGPU Types
      • Hopper H100 vGPU Types
      • Hopper H20 vGPU Types
    • Ada Lovelace Architecture vGPU Types
    • Ampere Architecture vGPU Types
      • Ampere Workstation and General Compute vGPU Types
      • Ampere A800 and AX800 vGPU Types
      • Ampere A100 vGPU Types
      • Ampere A30 vGPU Types
    • Turing Architecture vGPU Types

Troubleshooting

  • Overview
  • Glossary
  • Features
  • NVLink Multicast
Is this page helpful?

NVLink Multicast#

NVLink multicast is an NVSwitch fabric capability. Start with NVIDIA NVSwitch for fabric context, then use the support tables on this page.

Important

NVLink multicast support requires that unified memory is enabled. For more information about enabling unified memory, refer to the Enabling Unified Memory for a vGPU documentation, or Unified Virtual Memory (UVM) for in-repo coverage.

See also

NVLink Multicast (on NVSwitch) on the NVSwitch page.

vGPU Support for NVLink Multicast#

Only full-sized, time-sliced NVIDIA vGPU for Compute support NVLink multicast.

Table 61 NVIDIA NVLink Multicast Support on the NVIDIA Blackwell GPU Architecture#

Board

vGPU

NVIDIA HGX B300

NVIDIA B300X-279C

NVIDIA HGX B200

NVIDIA B200X-180C

Table 62 NVIDIA NVLink Multicast Support on the NVIDIA Hopper GPU Architecture#

Board

vGPU

NVIDIA HGX H800

NVIDIA H800XM-80C

NVIDIA HGX H200

NVIDIA H200X-141C

NVIDIA HGX H100

NVIDIA H100XM-80C

NVIDIA H20 HGX 141 GB

H20X-141C

HGX H20 96 GB

NVIDIA H20-96C

NVLink Multicast Limitations#

  • NVLink multicast is supported only with full-sized, time-sliced NVIDIA vGPU for Compute profiles. MIG-backed vGPU profiles and fractional time-sliced profiles are not supported.

  • Unified memory must be enabled for the vGPU. Refer to the Enabling Unified Memory for a vGPU documentation.

  • NVLink multicast requires NVSwitch-connected GPU boards (HGX platforms). PCIe-only GPU configurations do not support NVLink multicast.

  • Only the specific board and vGPU profile combinations listed in the support tables above are supported.

previous

NVIDIA NVSwitch

next

Scheduling Policies

On this page
  • vGPU Support for NVLink Multicast
  • NVLink Multicast Limitations
NVIDIA NVIDIA
Privacy Policy | Your Privacy Choices | Terms of Service | Accessibility | Corporate Policies | Product Security | Contact

Copyright © 2021-2026, NVIDIA Corporation.

Last updated on Aug 17, 2026.