A Comprehensive Survey on Cloud Data Mining (CDM) Frameworks and Algorithms

被引:17
|
作者
Barua, Hrishav Bakul [1 ,3 ]
Mondal, Kartick Chandra [2 ]
机构
[1] Embedded Syst & Robot Res Grp, TCS Res & Innovat Lab, Kolkata, India
[2] Jadavpur Univ, Dept Informat Technol, Sect 3, Kolkata 700106, W Bengal, India
[3] TCS Ecospace, TCS Res & Innovat Lab, Act Area 2, Kolkata 700156, W Bengal, India
关键词
Review; survey; taxonomy; framework; data mining; machine learning; distributed computing; cloud data mining (CDM); big data; big data analytics; data science; cloud computing; parallelism; graph mining; volume; velocity; variety; clustering; classification and association rule mining; BIG DATA; CLUSTERING ALGORITHMS; DATA ANALYTICS; MAPREDUCE; DBSCAN; CLASSIFICATION; PRIVACY; STORAGE; RISE;
D O I
10.1145/3349265
中图分类号
TP301 [理论、方法];
学科分类号
081202 ;
摘要
Data mining is used for finding meaningful information out of a vast expanse of data. With the advent of Big Data concept, data mining has come to much more prominence. Discovering knowledge out of a gigantic volume of data efficiently is a major concern as the resources are limited. Cloud computing plays a major role in such a situation. Cloud data mining fuses the applicability of classical data mining with the promises of cloud computing. This allows it to perform knowledge discovery out of huge volumes of data with efficiency. This article presents the existing frameworks, services, platforms, and algorithms for cloud data mining. The frameworks and platforms are compared among each other based on similarity, data mining task support, parallelism, distribution, streaming data processing support, fault tolerance, security, memory types, storage systems, and others. Similarly, the algorithms are grouped on the basis of parallelism type, scalability, streaming data mining support, and types of data managed. We have also provided taxonomies on the basis of data mining techniques such as clustering, classification, and association rule mining. We also have attempted to discuss and identify the major applications of cloud data mining. The various taxonomies for cloud data mining frameworks, platforms, and algorithms have been identified. This article aims at gaining better insight into the present research realm and directing the future research toward efficient cloud data mining in future cloud systems.
引用
收藏
页数:62
相关论文
共 50 条
  • [1] A comprehensive survey of data mining
    Gupta M.K.
    Chandra P.
    International Journal of Information Technology, 2020, 12 (4) : 1243 - 1257
  • [2] Issues in Data mining: A comprehensive survey
    Purwar, Archana
    Singh, Sandeep Kumar
    2014 IEEE INTERNATIONAL CONFERENCE ON COMPUTATIONAL INTELLIGENCE AND COMPUTING RESEARCH (IEEE ICCIC), 2014, : 657 - 662
  • [3] Big Data Security Survey on Frameworks and Algorithms
    Chandra, Sudipta
    Ray, Soumya
    Goswami, R. T.
    2017 7TH IEEE INTERNATIONAL ADVANCE COMPUTING CONFERENCE (IACC), 2017, : 48 - 54
  • [4] A Comprehensive Survey on Privacy Preservation Algorithms in Data Mining
    Kiran, Ajmeera
    Vasumathi, D.
    2017 IEEE INTERNATIONAL CONFERENCE ON COMPUTATIONAL INTELLIGENCE AND COMPUTING RESEARCH (ICCIC), 2017, : 1060 - 1066
  • [5] Data Mining Algorithms for Smart Cities: A Bibliometric Analysis
    Kousis, Anestis
    Tjortjis, Christos
    ALGORITHMS, 2021, 14 (08)
  • [6] A survey on parallel clustering algorithms for Big Data
    Zineb Dafir
    Yasmine Lamari
    Said Chah Slaoui
    Artificial Intelligence Review, 2021, 54 : 2411 - 2443
  • [7] Transplantation of Data Mining Algorithms to Cloud Computing Platform when Dealing Big Data
    Wang, Yong
    Zhao, Ya-Wei
    2014 INTERNATIONAL CONFERENCE ON CYBER-ENABLED DISTRIBUTED COMPUTING AND KNOWLEDGE DISCOVERY (CYBERC), 2014, : 175 - 178
  • [8] A survey on parallel clustering algorithms for Big Data
    Dafir, Zineb
    Lamari, Yasmine
    Slaoui, Said Chah
    ARTIFICIAL INTELLIGENCE REVIEW, 2021, 54 (04) : 2411 - 2443
  • [9] Big Data Security in Healthcare Survey on Frameworks and Algorithms
    Chandra, Sudipta
    Ray, Soumya
    Goswami, R. T.
    2017 7TH IEEE INTERNATIONAL ADVANCE COMPUTING CONFERENCE (IACC), 2017, : 89 - 94
  • [10] A Comprehensive Survey on Variants And Its Extensions Of Big Data In Cloud Environment
    Karthikeyan, P.
    Amudhavel, J.
    Abraham, A.
    Sathian, D.
    Raghav, R. S.
    Dhavachelvan, P.
    ICARCSET'15: PROCEEDINGS OF THE 2015 INTERNATIONAL CONFERENCE ON ADVANCED RESEARCH IN COMPUTER SCIENCE ENGINEERING & TECHNOLOGY (ICARCSET - 2015), 2015,