Robot: An Efficient Model For Big Data Storage Systems Based On Erasure Coding

被引:0
作者
Yin, Chao [1 ]
Wang, Jianzong [3 ]
Xie, Changsheng [1 ,2 ]
Wan, Jiguang [1 ]
Long, Changlin [1 ]
Bi, Wenjuan [1 ]
机构
[1] Huazhong Univ Sci & Technol, Sch Comp Sci & Technol, Wuhan, Peoples R China
[2] Wuhan Natl Lab Optoelect, Wuhan, Peoples R China
[3] NetEase Inc, Guangzhou, Guangdong, Peoples R China
来源
2013 IEEE INTERNATIONAL CONFERENCE ON BIG DATA | 2013年
基金
中国国家自然科学基金;
关键词
distributed file system; erasure coding; big data; robustness; availiabilty; cloud storage; MDS ARRAY CODES;
D O I
暂无
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
it is well-known that with the explosive growth of data, the age of big data has arrived. How to save huge amounts of data is of great importance to both industry and academia. This paper puts forward a solution based on coding technologies in big data system that store a lot of cold data. By studying existing coding technologies and big data systems, we can not only maintain the system's reliability, but also improve the security and the utilization of storage systems. Due to the remarkable reliability and space saving rate of coding technologies, importing coding schema in to big data systems becomes prerequisite. In our presented schema, the storage node is divided into several virtual nodes to keep load balancing. By setting up different virtual node storage groups for different codec server, we can ensure system availability. And by utilizing the parallel decoding computing of the node and the block of data, we can also reduce the system recovery time when data is corrupted. Additionally, different users set different coding parameters can improve the robustness of big data storage systems. We configure various data block m and calibration block k to improve the utilization rate in the quantitative experiments. The results shows that parallel decoding speed can rise up two times than the past serial decoding speed. The encoding efficiency with ICRS coding is 34.2% higher than using CRS and 56.5% more than using RS coding equally. The decoding rate by using ICRS is 18.1% higher than using CRS and 31.1% higher than using RS averagely.
引用
收藏
页数:6
相关论文
共 20 条
[1]  
[Anonymous], 2012, P USENIX ANN TECHN C
[2]  
[Anonymous], 2006, P 7 C OP SYST DES IM
[3]  
Blomer J., 1995, Tech. Rep. TR-95-048
[4]  
Feng GL, 2005, IEEE T COMPUT, V54, P1071, DOI 10.1109/TC.2005.150
[5]  
Ghemawat S., 2003, SOSP, P29
[6]  
Hafner J. L., 2006, DSN 06 INT C DEP SYS
[7]  
Hafner JL, 2005, USENIX ASSOCIATION PROCEEDINGS OF THE 4TH USENIX CONFERENCE ON FILE AND STORAGE TECHNOLOGIES, P211
[8]  
Li Yan, 11 USENIX C FIL STOR, P147
[9]   The evolution of storage systems [J].
Morris, RJT ;
Truskowski, BJ .
IBM SYSTEMS JOURNAL, 2003, 42 (02) :205-217
[10]  
Plank JS, 1997, SOFTWARE PRACT EXPER, V27, P995, DOI 10.1002/(SICI)1097-024X(199709)27:9<995::AID-SPE111>3.0.CO