RETRACTED ARTICLE: Detecting straggler MapReduce tasks in big data processing infrastructure by neural network

被引:0
作者
Amir Javadpour
Guojun Wang
Samira Rezaei
Kuan-Ching Li
机构
[1] Guangzhou University,School of Computer Science
[2] University of Groningen,Bernoulli Institute for Mathematics and Computer Science
[3] Providence University,Department of Computer Science and Information Engineering
来源
The Journal of Supercomputing | 2020年 / 76卷
关键词
Hadoop; Speculative execution; Straggler tasks; MapReduce; Artificial neural network;
D O I
暂无
中图分类号
学科分类号
摘要
Straggler task detection is one of the main challenges in applying MapReduce for parallelizing and distributing large-scale data processing. It is defined as detecting running tasks on weak nodes. Considering two stages in the Map phase (copy, combine) and three stages of Reduce (shuffle, sort and reduce), the total execution time is the total sum of the execution time of these five stages. Estimating the correct execution time in each stage that results in correct total execution time is the primary purpose of this paper. The proposed method is based on the application of a backpropagation neural network on the Hadoop for the detection of straggler tasks, to estimate the remaining execution time of tasks that is very important in straggler task detection. Results achieved have been compared with popular algorithms in this domain such as LATE, ESAMR and the real remaining time for WordCount and Sort benchmarks, and shown able to detect straggler tasks and estimate execution time accurately. Besides, it supports to accelerate task execution time.
引用
收藏
页码:6969 / 6993
页数:24
相关论文
共 53 条
[51]  
Wu AY(undefined)undefined undefined undefined undefined-undefined
[52]  
Peng Z(undefined)undefined undefined undefined undefined-undefined
[53]  
Wang G(undefined)undefined undefined undefined undefined-undefined