country_code
Skip to main content
Ctrl+K
NVSHMEM APIs - Home NVSHMEM APIs - Home

NVSHMEM APIs

NVSHMEM APIs - Home NVSHMEM APIs - Home

NVSHMEM APIs

Table of Contents

  • Introduction
  • Installing NVSHMEM
  • Using NVSHMEM
  • NVSHMEM and the CUDA Model
  • Using TMA with NVSHMEM
  • Memory Model
  • Execution Model
  • Library Constants
  • Library Handles
  • Environment Variables
  • NVSHMEM APIs
    • Overview of the APIs
    • Library Setup, Exit, and Query
    • Kernel Launch Routines
    • Memory Management
    • Queue Pair (QP) Specific APIs
    • Team Management
    • Remote Memory Access
    • Atomic Memory Operations
    • Signaling Operations
    • Collective Communication
    • Point-To-Point Synchronization
    • Memory Ordering
    • Language Bindings
      • Python Bindings (NVSHMEM4Py)
        • NVSHMEM4Py Overview
        • Initialization and Finalization
        • Memory Management
        • Interoperability
        • Collective Operations
        • Remote Memory Access (RMA)
        • Utility Functions for NVSHMEM4Py
        • Python Device APIs
  • Examples
    • Language Bindings Examples
      • NVSHMEM4Py Examples
  • Troubleshooting and FAQs
  • NVSHMEM SLA
  • Acknowledgements
  • Examples
  • Language Bindings Examples
Is this page helpful?

Language Bindings Examples#

Contents:

  • NVSHMEM4Py Examples
    • UID-Based Initialization Example
    • MPI Comm-Based Initialization Example
    • Torch.distributed ProcessGroup Initialization Example
    • Simple P2P Kernel Example
    • On-Stream Kernels Example
    • PyTorch and Triton Interoperability Example

This directory contains examples of using the NVSHMEM language bindings.

previous

Examples

next

NVSHMEM4Py Examples

NVIDIA NVIDIA
Privacy Policy | Your Privacy Choices | Terms of Service | Accessibility | Corporate Policies | Product Security | Contact

Copyright © 2022-2026, NVIDIA Corporation. All rights reserved..