Semi-supervised possibilistic c-means clustering algorithm based on feature weights for imbalanced data

被引:10
作者
Yu, Haiyan [1 ]
Xu, Xiaoyu [1 ]
Li, Honglei [1 ]
Wu, Yuting [1 ]
Lei, Bo [1 ]
机构
[1] Xian Univ Posts & Telecommun, Sch Telecommun & Informat Engn, Xian 710121, Peoples R China
基金
中国国家自然科学基金;
关键词
Clustering; Possibilistic c -means clustering (PCM); Semi; -supervised; Feature weight; Imbalanced data; Image segmentation; MAHALANOBIS DISTANCE; FUZZY; ENTROPY;
D O I
10.1016/j.knosys.2024.111388
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
The possibilistic c-means clustering (PCM) algorithm improves the robustness of fuzzy c-means clustering (FCM) to noise and outliers by releasing the probabilistic constraint of memberships. The semi-supervised possibilistic cmeans clustering (SSPCM) algorithm improves the clustering effect on datasets with imbalanced sizes by introducing a small amount of label information. However, the traditional semi-supervised algorithm still faces the problem of low utilization of supervision information for datasets with large differences in sample sizes. Moreover, the Euclidean distance, which treats features equally, cannot handle feature-imbalanced data. Therefore, this paper proposes a semi-supervised possibilistic c-means clustering algorithm based on feature weights (FW-SSPCM) by introducing the ideas of supervised centers. First, the algorithm introduces the supervised center into the objective function of the SSPCM to improve the utilization rate of supervision information and thus guide the center iteration of small clusters. Second, the feature weighting strategy is introduced in the objective function to adaptively assign feature weights according to the importance of different features in different clusters, thus improving the adaptability of the algorithm to feature-imbalanced datasets. In addition, to improve the robustness of the antinoise effect and retain additional image details, a new image segmentation algorithm based on FW-SSPCM and local information (LFW-SSPCM) is proposed by introducing local spatial information obtained by bilateral filtering. Finally, through clustering experiments on synthetic data, UCI datasets and on color images characteristic of multiple features, including imbalanced sizes, imbalanced features and strong noise injection, the clustering performances of the proposed FW-SSPCM and LFW-SSPCM proposed in this paper are significantly better than those of several related clustering algorithms.
引用
收藏
页数:37
相关论文
共 50 条
  • [21] A Performance Study of Probabilistic Possibilistic Fuzzy C-Means Clustering Algorithm
    Vijaya, J.
    Syed, Hussian
    ADVANCES IN COMPUTING AND DATA SCIENCES, PT I, 2021, 1440 : 431 - 442
  • [22] Sparse possibilistic c-means clustering with Lasso
    Yang, Miin-Shen
    Benjamin, Josephine B. M.
    PATTERN RECOGNITION, 2023, 138
  • [23] Improved possibilistic C-means clustering algorithms
    Zhang, JS
    Leung, YW
    IEEE TRANSACTIONS ON FUZZY SYSTEMS, 2004, 12 (02) : 209 - 217
  • [24] EM-IFCM: Fuzzy c-means clustering algorithm based on edge modification for imbalanced data
    Pu, Yue
    Yao, Wenbin
    Li, Xiaoyong
    INFORMATION SCIENCES, 2024, 659
  • [25] Context Data Clustering Based On Modified Fuzzy Possibilistic C-Means Algorithm for Efficient Context-Aware Computing Services
    Saad, Mohamed Fadhel
    Lee, Jongyoun
    Kwon, Ohbyung
    Alimi, Adel M.
    INFORMATION-AN INTERNATIONAL INTERDISCIPLINARY JOURNAL, 2011, 14 (09): : 3101 - 3111
  • [26] A feature-weighted suppressed possibilistic fuzzy c-means clustering algorithm and its application on color image segmentation
    Yu, Haiyan
    Jiang, Lerong
    Fan, Jiulun
    Xie, Shuang
    Lan, Rong
    EXPERT SYSTEMS WITH APPLICATIONS, 2024, 241
  • [27] Total-aware suppressed possibilistic c-means clustering
    Wu, Chengmao
    Xiao, Xue
    MEASUREMENT, 2023, 219
  • [28] Kernel possibilistic fuzzy c-means clustering algorithm based on morphological reconstruction and membership filtering
    Farooq, Anum
    Memon, Kashif Hussain
    FUZZY SETS AND SYSTEMS, 2024, 477
  • [29] On tolerant fuzzy c-means clustering and tolerant possibilistic clustering
    Yukihiro Hamasuna
    Yasunori Endo
    Sadaaki Miyamoto
    Soft Computing, 2010, 14 : 487 - 494
  • [30] A Semi-supervised Clustering Algorithm Based on Rough Reduction
    Lin, Liandong
    Qu, Wei
    Yu, Xiang
    CCDC 2009: 21ST CHINESE CONTROL AND DECISION CONFERENCE, VOLS 1-6, PROCEEDINGS, 2009, : 5427 - +