Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-09-20 | 0 | 0 | 1 | 1 | 2 |
| 2026-09-19 | 0 | 0 | 6 | 14 | 4 |
| 2026-09-18 | 2 | 0 | 5 | 10 | 5 |
| 2026-09-17 | 3 | 1 | 5 | 10 | 5 |
| 2026-09-16 | 0 | 2 | 3 | 4 | 3 |
| 2026-09-15 | 0 | 1 | 8 | 21 | 5 |
| 2026-09-14 | 2 | 5 | 12 | 29 | 8 |
| 2026-09-13 | 0 | 0 | 1 | 7 | 278 |
| 2026-09-12 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-11 | 0 | 0 | 0 | 0 | 0 |
| 2026-09-10 | 1 | 0 | 0 | 0 | 0 |
| 2026-09-09 | 1 | 0 | 0 | 0 | 0 |
| 2026-09-08 | 0 | 1 | 0 | 0 | 0 |
| 2026-09-07 | 0 | 0 | 0 | 0 | 0 |
Opened: 7
Closed: 9
Comments: 40
Events: 89
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#2181 [Question]: Does NCCL have any plan to provide QP counter metrics of GDAKI 's QPs? Opened 4 months ago | baymaxhuang | open | question | 5 | 0 | 15 hours ago |
#2332 [Issue]: `ncclCommRegister` fails with `CUDA_ERROR_NOT_FOUND` (500) on memory allocated with CUDA VMM. Opened 1 month ago | thearusable | open | No labels | 4 | 0 | 22 hours ago |
#2417 [RFE]: Add optional receive queues to the GDAKI backend Opened 3 days ago | mkbk-with-circle | open | No labels | 3 | 0 | 1 day ago |
#2416 [RFE]:Pluggable external device-memory allocator: hand all non-registration buffer classes to a framework-supplied allocator (e.g. PyTorch's caching allocator) Opened 4 days ago | zjjott | open | enhancement | 2 | 0 | 2 days ago |
#2420 NVCC_GENCODE_LDMC_FP8 device-code gencode ignores the caller's NVCC_GENCODE target-arch list, breaking single-arch builds Opened 3 days ago | zbrad | open | No labels | 1 | 0 | 2 days ago |
#2421 ncclAlltoAllConfig does not honor per-call maxCTAs Opened 2 days ago | ngoyal2707 | open | No labels | 1 | 0 | 2 days ago |
#2033 [Issue]: CPU affinity is not restored at the end of `initTransportsRank()` (introduced in 2.29.2-1) Opened 7 months ago | shaitan | closed - completed | No labels | 3 | 1 | 2 days ago |
#1270 Provenance of NVTX headers in NCCL Opened 2 years ago | Artem-B | open | No labels | 14 | 1 | 2 days ago |
#2111 [RFC] nccl-ep: support Attention–MoE disaggregation Opened 5 months ago | kwen2501 | open | No labels | 1 | 1 | 2 days ago |
#1026 half precision reduction accumulation in fp32? Opened 3 years ago | stas00 | closed - not_planned | No labels | 16 | 0 | 3 days ago |
#2308 [Announcement] NCCL EP and NCCL M2N are moving to NVIDIA/nccl-extensions Opened 2 months ago | kwen2501 | open | Community Discussion | 2 | 4 | 3 days ago |
#2418 [Issue]: Dual RTX PRO 6000 Blackwell: NCCL 2.26.2 illegal memory access for >=512 KiB collectives, not reproducible with 2.31.2 Opened 3 days ago | AlbertLee98 | open | No labels | 1 | 0 | 3 days ago |
#2372 [Question]: NCCL GIN GDAKI (type 3) hangs during Pure GIN AlltoAll, PROXY (type 2) works Opened 25 days ago | HPC4AI | open | question | 1 | 0 | 3 days ago |
#2397 [Issue]: ncclCommShrink leads to IB transport errors and hang when excluded ranks exit without calling shrink, and also affects the allreduce operations after the shrink. Opened 11 days ago | perpyoke | closed - completed | No labels | 6 | 0 | 3 days ago |
#2360 [Bug] Hopper NVLS BF16 results depend on multicast allocation history and differ across H800 hosts with identical software Opened 1 month ago | wqshr12345 | open | No labels | 7 | 0 | 4 days ago |
#2325 [Issue]: After the rollback of `PciInfo` the `nvmlDeviceGetPciInfo_v3` is used with `v2` struct causing stack buffer overflow. Opened 2 months ago | thearusable | closed - completed | No labels | 4 | 0 | 4 days ago |
#2226 [Question]: [GIN] GPI vs GDAKI: who selects the QP, and what NIC support does GPI need? Opened 3 months ago | QizhouZhang97 | closed - completed | question | 3 | 0 | 4 days ago |
#2408 NCCL_CHUNK_SIZE / NCCL_P2P_NET_CHUNKSIZE >= 1MiB gives incorrect results and Xid 31 GPU MMU faults (H100, 8x400G RoCE) Opened 7 days ago | Ndministrator | open | No labels | 3 | 0 | 4 days ago |
#2409 Low AllReduce bandwidth on 2x8x H100 with 8x 400G RoCE: per GPU<->GPU connection capped at ~13 GB/s while perftest reaches 392 Gb/s line rate Opened 7 days ago | Ndministrator | closed - completed | No labels | 4 | 0 | 4 days ago |
#2160 NCCL_TOPO_XML_MAX_NODES=256 limit hit during intra-node XML fusion on 32-NIC hosts (AWS p5.48xlarge) Opened 5 months ago | dmvevents | open | No labels | 4 | 0 | 5 days ago |
#2367 [Issue]: dlopen() of a truncated plugin .so causes SIGBUS crash instead of graceful fallback (NCCL_PROFILER_PLUGIN / NET / TUNER / ENV) Opened 28 days ago | AdamSabry1233 | closed - not_planned | No labels | 8 | 0 | 6 days ago |
#2345 [Issue]: Ring and tree both create cross-rail inter-node links, violating rail isolation and despite `NCCL_CROSS_NIC=0` Opened 1 month ago | StevenSong | open | No labels | 6 | 0 | 6 days ago |
#2379 Probabilistic multi-communicator init hang in the cuMem UDS fd exchange on single-node PCIe (no NVLink); NCCL_CUMEM_ENABLE=0 isolates it, 20/20 vs 5/9 A/B Opened 22 days ago | tomylin890 | open | No labels | 3 | 0 | 6 days ago |
#2355 [Issue]: Profiler plugin overhead regression after upgrading to NCCL 2.31.2 Opened 1 month ago | hzhengBd | closed - completed | No labels | 4 | 0 | 6 days ago |
#2238 [Issue]: Performance and functional issues when using ncclPutSignal/ncclWaitSignal in Pipeline Parallelism Opened 3 months ago | DAMI211 | open | No labels | 13 | 0 | 6 days ago |