会议3_利用现代 GPU 集群基于 GPU 的高效压缩方案加速 MPI ALLREDUCE 通信.pdf

编号:171248 PDF 21页 2.60MB 下载积分:VIP专享
下载报告请您先登录!

会议3_利用现代 GPU 集群基于 GPU 的高效压缩方案加速 MPI ALLREDUCE 通信.pdf

1、ACCELERATING MPI ALLREDUCECOMMUNICATION WITH EFFICIENT GPU-BASED COMPRESSION SCHEMES ON MODERN GPU CLUSTERS2024 OFA Virtual WorkshopHari Subramoni and Qinghua ZhouThe Ohio State UniversityEmail:subramoni.1,Zhou.2595osu.eduPRESENTATION OVERVIEWIntroduction&MotivationDesign ApproachesRing AllReduce wi

2、th Collective-level Online CompressionRecursive-Doubling AllReduce with Collective-level Online CompressionPerformance EvaluationBenchmark-level evaluationApplication-level evaluationConclusion and Future Plan2 OpenFabrics AllianceINTRODUCTION AND MOTIVATIONAllReduce is a communication collective op

3、eration that is commonly used in HPC applications as well as distributed DL training.Existing AllReduce algorithms for transferring large GPU data still suffer from poor performance due to the limited interconnect bandwidth of networksNaive point-to-point compression for each data transmission may i

4、ntroduce redundant compression/decompression operations and hinder non-blocking send/receive operationsHow to co-design and optimize the GPU-based compression at the collective-level along with the communication patterns of advanced AllReduce algorithms?We propose two design approaches along with th

5、ese directions.Ring AllReduce with Collective-level Online CompressionRecursive-Doubling AllReduce with Collective-level Online Compression3 OpenFabrics AlliancePRESENTATION OVERVIEWIntroduction&MotivationDesign ApproachesRing AllReduce with Collective-level Online CompressionRecursive-Doubling AllR

6、educe with Collective-level Online CompressionPerformance EvaluationBenchmark-level evaluationApplication-level evaluationConclusion and Future Plan4 OpenFabrics AllianceDESIGN APPROACHESRing and Recursive-Doubling MPI_AllReducewith Collective-level Online CompressionCompression can reduce the data

友情提示

1、下载报告失败解决办法
2、PDF文件下载后,可能会被浏览器默认打开,此种情况可以点击浏览器菜单,保存网页到桌面,就可以正常下载了。
3、本站不支持迅雷下载,请使用电脑自带的IE浏览器,或者360浏览器、谷歌浏览器下载即可。
4、本站报告下载后的文档和图纸-无水印,预览文档经过压缩,下载后原文更清晰。

本文(会议3_利用现代 GPU 集群基于 GPU 的高效压缩方案加速 MPI ALLREDUCE 通信.pdf)为本站 (Chriswl) 主动上传,三个皮匠报告文库仅提供信息存储空间,仅对用户上传内容的表现方式做保护处理,对上载内容本身不做任何修改或编辑。 若此文所含内容侵犯了您的版权或隐私,请立即通知三个皮匠报告文库(点击联系客服),我们立即给予删除!

温馨提示:如果因为网速或其他原因下载失败请重新下载,重复下载不扣分。
客服
商务合作
小程序
服务号
折叠