-
Notifications
You must be signed in to change notification settings - Fork 103
Pull requests: microsoft/mscclpp
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
Proxy FIFO latency improvement
#864
opened Aug 7, 2026 by
Changho Hwang (chhwang)
Contributor
•
1/2
Loading…
Add NVLS switch-channel cross-rank barrier api
#863
opened Aug 7, 2026 by
RJ Souza (Empyreus)
Contributor
Loading…
Add device code review skill and refresh CI
#862
opened Aug 6, 2026 by
Binyang Li (Binyang2014)
Contributor
Loading…
ep(bench): add mscclpp high-throughput (rank-major) backend to the unified Python EP benchmark
#860
opened Aug 3, 2026 by
Qinghua Zhou (seagater)
Contributor
Loading…
ep(bench): unified in-process Python EP benchmark — review-comment follow-ups
#858
opened Jul 30, 2026 by
Qinghua Zhou (seagater)
Contributor
Loading…
Add bulk asynchronous copy primitives
#853
opened Jul 24, 2026 by
Changho Hwang (chhwang)
Contributor
Loading…
Add switch-channel NVLS cross-rank barrier
#843
opened Jul 16, 2026 by
RJ Souza (Empyreus)
Contributor
Loading…
Extending DSL Executor CI tests for multi node envinronment
#812
opened May 22, 2026 by
Caio Rocha (caiocbr)
Contributor
Loading…
Comprehensive GDRCopy version check
#805
opened May 15, 2026 by
Changho Hwang (chhwang)
Contributor
Loading…
Introduce GPUBufferPool for symmetric memory support
#798
opened May 6, 2026 by
Binyang Li (Binyang2014)
Contributor
Loading…
Migrate MoE dispatch / combine primitives into MSCCL++
#796
opened May 4, 2026 by
Qinghua Zhou (seagater)
Contributor
Loading…
Add FP8 conversion unit tests
#789
opened Apr 17, 2026 by
Binyang Li (Binyang2014)
Contributor
Loading…
Add
accumulate method for PortChannel
#784
opened Apr 13, 2026 by
Changho Hwang (chhwang)
Contributor
•
2/2
Loading…
TorchComms Integration for MSCCL++
#771
opened Apr 5, 2026 by
Michael Beebe (michael-beebe)
Loading…
Unique QP per channel and env-controlled GID index for executor
#762
opened Mar 9, 2026 by
Binyang Li (Binyang2014)
Contributor
•
Draft
Align the data element counts for Allreduce nvls kernel
#697
opened Dec 1, 2025 by
Qinghua Zhou (seagater)
Contributor
Loading…
Alltoallv with dynamic execution plan template
#629
opened Sep 9, 2025 by
Qinghua Zhou (seagater)
Contributor
Loading…
Track Intra Rank Data Dependency and Automatically Sync Thread in the same Rank
#627
opened Sep 8, 2025 by
Caio Rocha (caiocbr)
Contributor
•
Draft
Fifo test with data transfer between multiple GPUs
#596
opened Aug 5, 2025 by
Qinghua Zhou (seagater)
Contributor
Loading…
Try to extend the connection
#578
opened Jul 23, 2025 by
Binyang Li (Binyang2014)
Contributor
•
Draft
Fifo test with multiple gpus
#573
opened Jul 22, 2025 by
Qinghua Zhou (seagater)
Contributor
Loading…
Add indirect connection support
#571
opened Jul 19, 2025 by
Zhaoyang (Cyrus) Hong (hfhongzy)
Loading…
Exchange recvbuff at the beginning of the CUDA kernel call.
#464
opened Feb 19, 2025 by
SreevatsaAnantharamu
Contributor
Loading…
Previous Next
ProTip!
Add no:assignee to see everything that’s not assigned.