Protocol

NCCL

NVIDIA Collective Communications Library

Lab ready Full lab

NVIDIA Collective Communications Library implementing allreduce, broadcast, and related GPU collectives over NVLink, PCIe, and RDMA. Topology detection and ring/tree algorithms dominate training performance at scale. It is the default collective layer for many PyTorch/TensorFlow multi-GPU jobs.

Domains

Tags

collectivesgpuallreducetraining

Related protocols

ProtoLab · signal bench for protocols

Privacy Shane Code

Signal list

Lab notes, not a campaign dump

Private ProtoLab list only. Confirm the address and you can leave any time.

Check email to confirm. Privacy.