$npx skillfedfor your agent

nvidia-nvshmem-cu12

NVSHMEM creates a global address space that provides efficient and scalable communication for NVIDIA GPU clusters.

With conditionsPyPI Software DevelopmentReleased Jul 20267.6M downloads / moPlatform wheel

Decision gist · record as of 2026-08-14

platform wheels — nvidia_nvshmem_cu12-3.7.2-py3-none-manylinux2014_aarch64.manylinux_2_17_aarch64.whl · nvidia_nvshmem_cu12-3.7.2-py3-none-manylinux2014_x86_64.manylinux_2_17_x86_64.whl
v3.7.2 · released 2026-07-17 · Python >=3 · 1 runtime deps: nvidia-cuda-cccl-cu12

Yes, if you are building multi-GPU CUDA applications on Linux and need efficient inter-GPU communication. The package is actively maintained, has no known vulnerabilities, and targets a real use case in distributed GPU computing. However, verify the unclear license terms and confirm Windows support is not actually available despite classifier claims. Medium install friction is acceptable for a specialized GPU library.AI-flagged interpretation of the facts on this page — verify before relying

Before you install

  • Requires NVIDIA CUDA 12.x runtime, NVIDIA GPUs, and Linux (aarch64 or x86_64); not available on Windows or macOS despite classifier claims.
  • Medium install friction due to platform-specific wheels (aarch64 and x86_64 Linux only).
  • Active maintenance with a recent release, though no public repository or commit history is available to verify ongoing development velocity.

License · maintenance · safety

(unclear) — License terms are unclear—no SPDX identifier or raw license text is published. Verify licensing compatibility with your project before production use.

last release 2026-07-17 (28 days)

0 known vulnerabilities (OSV.dev, 2026-08-14) · 7,604,088 downloads/mo, #1,715 on PyPI

Verify before relying

pip install nvidia-nvshmem-cu12
import nvshmem
# Initialize NVSHMEM for multi-GPU communication
nvshmem.init()
  • Actual Windows support status—classifiers list Windows but wheels are Linux-only.
  • Whether public documentation exists beyond the NVIDIA CUDA Zone homepage.
  • Specific CUDA 12 minor version requirements and GPU compute capability minimums.
  • Whether the package works standalone or requires additional NVIDIA libraries beyond nvidia-cuda-cccl-cu12.
Same gist for agents: .md · .json

What it is and what it does

NVSHMEM is NVIDIA's implementation of the OpenSHMEM parallel programming standard, adapted for GPU clusters. It creates a unified memory address space spanning multiple GPUs, allowing kernels and CPU code to read and write remote GPU memory with fine-grained control. The package is built on top of CUDA 12 and is intended for researchers and engineers building distributed machine learning and scientific computing workloads that need efficient inter-GPU communication.

The package is in Beta status and actively maintained. It targets modern Python versions (3.5 through 3.11) and runs on Linux (both aarch64 and x86_64 architectures). Installation requires the nvidia-cuda-cccl-cu12 runtime dependency, and the platform-specific wheels mean you must match your system architecture exactly.

Use it for

  • Implement collective operations (allreduce, broadcast) across GPUs in a cluster without explicit message passing.
  • Build distributed training loops where multiple GPUs access shared parameter buffers via global memory semantics.
  • Develop scientific simulations that require fine-grained synchronization and data movement between GPU memories.
  • Optimize communication in multi-GPU inference pipelines by leveraging GPU-initiated remote memory access.

Worth the install?

AI-flagged interpretation of the facts on this page. Verify before relying on it.

With conditions

Yes, if you are building multi-GPU CUDA applications on Linux and need efficient inter-GPU communication.

The package is actively maintained, has no known vulnerabilities, and targets a real use case in distributed GPU computing. However, verify the unclear license terms and confirm Windows support is not actually available despite classifier claims. Medium install friction is acceptable for a specialized GPU library.

Install

nvidia-nvshmem-cu12 on PyPI

Before you install

Medium install friction due to platform-specific wheels (aarch64 and x86_64 Linux only). Active maintenance with a recent release, though no public repository or commit history is available to verify ongoing development velocity.

Requires NVIDIA CUDA 12.x runtime, NVIDIA GPUs, and Linux (aarch64 or x86_64); not available on Windows or macOS despite classifier claims.

License in practice

License terms are unclear—no SPDX identifier or raw license text is published. Verify licensing compatibility with your project before production use.

Quickstart

pip install nvidia-nvshmem-cu12
import nvshmem
# Initialize NVSHMEM for multi-GPU communication
nvshmem.init()

Verify before relying

  • Actual Windows support status—classifiers list Windows but wheels are Linux-only.
  • Whether public documentation exists beyond the NVIDIA CUDA Zone homepage.
  • Specific CUDA 12 minor version requirements and GPU compute capability minimums.
  • Whether the package works standalone or requires additional NVIDIA libraries beyond nvidia-cuda-cccl-cu12.

Package facts

LicenseNot declared unclear
Python supportSupports the current Python release >=3
Install frictionMedium. Platform-specific wheel
Runtime dependencies
1 package
nvidia-cuda-cccl-cu12
MaintenanceActively maintained 28 days since the last release
First released
Downloads7,604,088 / month, #1,715 on PyPI 30-day window, as of 2026-08-14
Known vulnerabilitiesNone known OSV.dev, checked 2026-08-14
Classifiers
Development Status :: 4 - BetaIntended Audience :: DevelopersIntended Audience :: EducationIntended Audience :: Science/ResearchNatural Language :: EnglishOperating System :: Microsoft :: WindowsOperating System :: POSIX :: LinuxProgramming Language :: Python :: 3Programming Language :: Python :: 3 :: OnlyProgramming Language :: Python :: 3.10Programming Language :: Python :: 3.11Programming Language :: Python :: 3.5Programming Language :: Python :: 3.6Programming Language :: Python :: 3.7Programming Language :: Python :: 3.8Programming Language :: Python :: 3.9Topic :: Scientific/EngineeringTopic :: Scientific/Engineering :: Artificial IntelligenceTopic :: Scientific/Engineering :: MathematicsTopic :: Software DevelopmentTopic :: Software Development :: Libraries

Evidence: nvidia_nvshmem_cu12-3.7.2-py3-none-manylinux2014_aarch64.manylinux_2_17_aarch64.whl; nvidia_nvshmem_cu12-3.7.2-py3-none-manylinux2014_x86_64.manylinux_2_17_x86_64.whl

Tags

Capabilities
gpu cluster communicationnvidia nvshmemmulti-gpu shared memoryopenshmem gpucuda inter-gpu communicationdistributed gpu memorygpu collective operations
Topics
gpu-computingdistributed-systemscuda
PyPI keywords
cudanvidiaruntimemachine learningdeep learning

Let your AI agent find packages like this

Example. Real query, live index.

You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.

wish › “gpu cluster communication”

Give your agent the search over MCP, or paste the wish link into any chat.

More Software Development packages

typing-extensions Worth it
PyPI · Software Development · released Jul 2026

Provides backported and experimental type hints for Python 3.9+, allowing use of newer typing features on older Python versions and enabling early experimentation with type system PEPs before they enter the standard library.

PSF-2.0pure Python · 3.9+
1.9Bdownloads / mo
numpy Worth it
PyPI · Software Development · released Aug 2026

NumPy provides an N-dimensional array object and a comprehensive suite of mathematical, linear algebra, Fourier transform, and random number functions for scientific computing in Python.

BSD-3-Clause AND 0BSD AND MIT AND Zlib AND CC0-1.0compiled wheel · 3.12+
1.1Bdownloads / mo
fastapi Worth it
PyPI · Software Development · released Jul 2026

FastAPI is a Python web framework for building REST APIs using type hints, with automatic request validation, serialization, and interactive API documentation.

MITpure Python · 3.10+
568.6Mdownloads / mo
annotated-doc With conditions
PyPI · Software Development · released Jul 2026

Provides a way to document function parameters, class attributes, return types, and variables inline using Python's `Annotated` type hint syntax instead of traditional docstrings.

MITpure Python · 3.9+
456.2Mdownloads / mo
typer Worth it
PyPI · Software Development · released Aug 2026

Typer builds command-line applications from Python functions using type hints, automatically generating help text, argument parsing, and shell completion.

Install it if you are building CLIs in Python.

MITpure Python · 3.10+
369.3Mdownloads / mo
distlib With conditions
PyPI · Software Development · released Jun 2026

Distlib provides low-level packaging utilities for building, distributing, and managing Python software—including metadata handling, version specifiers, wheel support, script installation, and dependency resolution.

permissive licensepure Python
323.3Mdownloads / mo

See also nvidia-nvshmem-cu13 · nvshmem4py-cu13 · nvidia-nccl-cu13 · nvidia-nccl-cu12 · nvidia-cuda-cccl-cu12 · nccl4py · nvidia-nccl-cu11 · nvidia-cusparse-cu12 · nvgpu · umf

Further reading