A Comprehensive Survey on Cloud Data Mining (CDM) Frameworks and Algorithms

被引:17
作者
Barua, Hrishav Bakul [1 ,3 ]
Mondal, Kartick Chandra [2 ]
机构
[1] Embedded Syst & Robot Res Grp, TCS Res & Innovat Lab, Kolkata, India
[2] Jadavpur Univ, Dept Informat Technol, Sect 3, Kolkata 700106, W Bengal, India
[3] TCS Ecospace, TCS Res & Innovat Lab, Act Area 2, Kolkata 700156, W Bengal, India
关键词
Review; survey; taxonomy; framework; data mining; machine learning; distributed computing; cloud data mining (CDM); big data; big data analytics; data science; cloud computing; parallelism; graph mining; volume; velocity; variety; clustering; classification and association rule mining; BIG DATA; CLUSTERING ALGORITHMS; DATA ANALYTICS; MAPREDUCE; DBSCAN; CLASSIFICATION; PRIVACY; STORAGE; RISE;
D O I
10.1145/3349265
中图分类号
TP301 [理论、方法];
学科分类号
081202 ;
摘要
Data mining is used for finding meaningful information out of a vast expanse of data. With the advent of Big Data concept, data mining has come to much more prominence. Discovering knowledge out of a gigantic volume of data efficiently is a major concern as the resources are limited. Cloud computing plays a major role in such a situation. Cloud data mining fuses the applicability of classical data mining with the promises of cloud computing. This allows it to perform knowledge discovery out of huge volumes of data with efficiency. This article presents the existing frameworks, services, platforms, and algorithms for cloud data mining. The frameworks and platforms are compared among each other based on similarity, data mining task support, parallelism, distribution, streaming data processing support, fault tolerance, security, memory types, storage systems, and others. Similarly, the algorithms are grouped on the basis of parallelism type, scalability, streaming data mining support, and types of data managed. We have also provided taxonomies on the basis of data mining techniques such as clustering, classification, and association rule mining. We also have attempted to discuss and identify the major applications of cloud data mining. The various taxonomies for cloud data mining frameworks, platforms, and algorithms have been identified. This article aims at gaining better insight into the present research realm and directing the future research toward efficient cloud data mining in future cloud systems.
引用
收藏
页数:62
相关论文
共 50 条
  • [41] Mining of Classification Patterns in Clinical Data through Data Mining Algorithms
    Jacob, Shomona Gracia
    Ramani, R. Geetha
    PROCEEDINGS OF THE 2012 INTERNATIONAL CONFERENCE ON ADVANCES IN COMPUTING, COMMUNICATIONS AND INFORMATICS (ICACCI'12), 2012, : 997 - 1003
  • [42] Systematic survey of big data and data mining in internet of things
    Shadroo, Shabnam
    Rahmani, Amir Masoud
    COMPUTER NETWORKS, 2018, 139 : 19 - 47
  • [43] A predictive approach to task scheduling for Big Data in Cloud environments using classification algorithms
    Vashishth, Vidushi
    Chhabra, Anshuman
    Sood, Apoorvi
    PROCEEDINGS OF THE 7TH INTERNATIONAL CONFERENCE ON CLOUD COMPUTING, DATA SCIENCE AND ENGINEERING (CONFLUENCE 2017), 2017, : 188 - 192
  • [44] A Survey - Data Mining Frameworks in Credit Card Processing
    Wongchinsri, Pornwatthana
    Kuratach, Werasak
    2016 13TH INTERNATIONAL CONFERENCE ON ELECTRICAL ENGINEERING/ELECTRONICS, COMPUTER, TELECOMMUNICATIONS AND INFORMATION TECHNOLOGY (ECTI-CON), 2016,
  • [45] A Comprehensive Survey and Open Challenges of Mining Bigdata
    Tidke, Bharat
    Mehta, Rupa
    Dhanani, Jenish
    INFORMATION AND COMMUNICATION TECHNOLOGY FOR INTELLIGENT SYSTEMS (ICTIS 2017) - VOL 1, 2018, 83 : 441 - 448
  • [46] Research on parallel data processing of data mining platform in the background of cloud computing
    Bu, Lingrui
    Zhang, Hui
    Xing, Haiyan
    Wu, Lijun
    JOURNAL OF INTELLIGENT SYSTEMS, 2021, 30 (01) : 479 - 486
  • [47] An experimental survey on big data frameworks
    Inoubli, Wissem
    Aridhi, Sabeur
    Mezni, Haithem
    Maddouri, Mondher
    Nguifo, Engelbert Mephu
    FUTURE GENERATION COMPUTER SYSTEMS-THE INTERNATIONAL JOURNAL OF ESCIENCE, 2018, 86 : 546 - 564
  • [48] Big Data and Cloud: A Survey
    Sangeetha, K. S.
    Prakash, P.
    ARTIFICIAL INTELLIGENCE AND EVOLUTIONARY ALGORITHMS IN ENGINEERING SYSTEMS, VOL 2, 2015, 325 : 773 - 778
  • [49] DATA MINING ON THE CANDELA CLOUD PLATFORM
    Yao, Wei
    Dumitru, Corneliu Octavian
    Lorenzo, Jose
    Datcu, Mihai
    IGARSS 2020 - 2020 IEEE INTERNATIONAL GEOSCIENCE AND REMOTE SENSING SYMPOSIUM, 2020, : 6945 - 6948
  • [50] MobSafe: Cloud Computing Based Forensic Analysis for Massive Mobile Applications Using Data Mining
    Xu, Jianlin
    Yu, Yifan
    Chen, Zhen
    Cao, Bin
    Dong, Wenyu
    Guo, Yu
    Cao, Junwei
    TSINGHUA SCIENCE AND TECHNOLOGY, 2013, 18 (04) : 418 - 427