No AI summary available for this article.
Why It Matters
Distributed inference depends on GPU collective communication that must keep pace with evolving hardware and specialized workloads.
Provenance
Discovered via ArXiv and published by ArXiv.
Key Claims
Original description
Distributed inference depends on GPU collective communication that must keep pace with evolving hardware and specialized workloads. However, existing collective implementations often couple semantics, orchestration (where and when data moves), and the datapath (how data moves). This coupling makes it costly to adopt new hardware mechanisms and customize communication for applications. We present Purlin, a scale-up communication framework that separates these concerns. At the top of Purlin, we specify collectives as a naming of an input and output layout and a copy or reduction operation. In th...
Discovered via ArXiv
Research papers and preprints from arXiv.
Publisher: arxiv.org
ID: http://arxiv.org/abs/2609.36954v1 · Indexed about 2 hours ago