Characterizing the Scalability of Graph Convolutional Networks on Intel^® PIUMA

Matthew Joseph Adiletta,Jesmin Jahan Tithi,Emmanouil-Ioannis Farsarakis,Gerasimos Gerogiannis,Robert Adolf, Robert Benke,Sidharth Kashyap,Samuel Hsia,Kartik Lakhotia,Fabrizio Petrini,Gu-Yeon Wei,David Brooks

2023 IEEE International Symposium on Performance Analysis of Systems and Software (ISPASS)（2023）

引用 0|浏览3

暂无评分

摘要

Large-scale Graph Convolutional Network (GCN) inference on traditional CPU/GPU systems is challenging due to a large memory footprint, sparse computational patterns, and irregular memory accesses with poor locality. Intel’s Programmable Integrated Unffied Memory Architecture (PIUMA) is designed to address these challenges for graph analytics. In this paper, a detailed characterization of GCNs is presented using the Open-Graph Benchmark (OGB) datasets to determine the viability of PIUMA as a potential solution to GCN scalability. First, the extent of sparse matrix dense matrix multiplication (SpMM) as a performance driver for GCN on CPU and GPU is explored, offering a methodology for predicting GCN behavior as a function of dataset characteristics. Second, an SpMM kernel optimized for PIUMA is described and investigated for sensitivity to system parameters including memory bandwidth, latency, and thread count. SpMM scalability on PIUMA is demonstrated, while the scalability limitations of a Xeon-optimized SpMM implementation are discussed. Finally, GCN performance is compared on PIUMA versus a Xeon CPU system and Ampere GPU system, showing impressive results on PIUMA for largescale datasets.

查看译文

关键词

Graph Convolution,SpMM,Memory Bandwidth Scaling,Latency Sensitivity,PIUMA,GCN

AI 理解论文

溯源树

样例

生成溯源树，研究论文发展脉络

Chat Paper

正在生成论文摘要

Characterizing the Scalability of Graph Convolutional Networks on Intel® PIUMA

Characterizing the Scalability of Graph Convolutional Networks on Intel^® PIUMA