Embedding GPU Computations in Hadoop

被引:4
作者
Zhu, Jie [1 ]
Jiang, Hai [1 ]
Li, Juanjuan [1 ]
Hardesty, Erikson [1 ]
Li, Kuan-Ching [2 ]
Li, Zhongwen [3 ]
机构
[1] Arkansas State Univ, Dept Comp Sci, State Univ, AR 72467 USA
[2] Providence Univ, Dept Comp Sci & Informat Engr, Taichung, Taiwan
[3] Chengdu Univ, Coll Informat Sci & Technol, Chengdu, Sichuan, Peoples R China
关键词
Hadoop; MapReduce; GPU; CUDA;
D O I
10.2991/ijndc.2014.2.4.2
中图分类号
TP31 [计算机软件];
学科分类号
081202 ; 0835 ;
摘要
As the size of high performance applications increases, four major challenges including heterogeneity, programmability, fault resilience, and energy efficiency have arisen in the underlying distributed systems. To tackle with all of them without sacrificing performance, traditional approaches in resource utilization, task scheduling and programming paradigm should be reconsidered. While Hadoop has handled data-intensive applications well in Clouds, GPU has demonstrated its acceleration effectiveness for computation-intensive ones. This paper addresses the approaches for Hadoop to exploiting both CPU and GPU resources effectively to handle aforementioned challenges. Hadoop schedules MapReduce's Map and Reduce functions across multiple different computing nodes through Java, whereas CUDA code helps accelerate local computations further on attached GPUs. All available heterogeneous computational power will be utilized. MapReduce in Hadoop eases the programming task by hiding communication and scheduling details. Hadoop Distributed File System will help achieve data-level fault resilience. GPU's energy efficiency characteristics help reduce the power consumption of the whole system. To utilize GPU in Hadoop, four approaches including Jcuda, JNI, Hadoop Streaming, and Hadoop Pipes, have been accomplished and analyzed. Experimental results have demonstrated and compared their effectiveness.
引用
收藏
页码:211 / 220
页数:10
相关论文
共 11 条
[1]  
Catanzaro B., 2008, P WORKSH SOFTW TOOLS
[2]  
Chen Linchuan, 2012, P INT C HIGH PERF CO
[3]  
Chen Linchuan, 2012, P HPDC
[4]  
Dean J., 2004, P 6 C S OP SYST DES
[5]  
Dean Jeffrey, COMMUNICATIONS ACM, P107
[6]  
Grossman M., 2013, P HDPIC 2013
[7]   Mars: A MapReduce Framework on Graphics Processors [J].
He, Bingsheng ;
Fang, Wenbin ;
Luo, Qiong ;
Govindaraju, Naga K. ;
Wang, Tuyong .
PACT'08: PROCEEDINGS OF THE SEVENTEENTH INTERNATIONAL CONFERENCE ON PARALLEL ARCHITECTURES AND COMPILATION TECHNIQUES, 2008, :260-269
[8]  
Okur S., HADOOP APARAPI MAKIN
[9]  
Stuart J. A., 2010, P 1 INT WORKSH MAPRE
[10]  
Yan Y., 2009, P EUR C SER AUG