כתבה
arXiv cs.AI ·
Purlin: פרידה בין התזמון לדרך הנתונים של קבוצות
Purlin: Separating Orchestration from the Datapath of Collectives
Purlin היא פרקטיקה של תקשורת שמפרידה בין התזמון לדרך הנתונים של קבוצות. היא מאפשרת לאדפט חדשות מכשירי GPU ולאפשר תקשורת מותאמת ליישומים. Purlin נבחנה על A100, H200, ו-B200 GPUs והציגה עד 5.14x עלייה במהירות ו-4.50x עלייה בתפוקה.
תקציר מקורי באנגליתarXiv:2609.36954v1 Announce Type: cross Abstract: Distributed inference depends on GPU collective communication that must keep pace with evolving hardware and specialized workloads. However, existing collective implementations often couple semantics, orchestration (where and when data moves), and the datapath (how data moves). This coupling makes it costly to adopt new hardware mechanisms and customize communication for applications. We present Purlin, a scale-up communication framework that separates these concerns. At the top of Purlin, we specify collectives as a naming of an input and output layout and a copy or reduction operation. In the middle, we introduce a shared orchestration protocol, Stage, Notify, And Consume (SNAC), which derives coordination from these specifications. Below
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית