Online cost-sensitive neural network classifiers for non-stationary and imbalanced data streams

被引:0
作者
Adel Ghazikhani
Reza Monsefi
Hadi Sadoghi Yazdi
机构
[1] Ferdowsi University of Mashhad,Computer Engineering Department
来源
Neural Computing and Applications | 2013年 / 23卷
关键词
Data stream classification; One-layer neural network; Concept drift; Imbalanced data; Online classification; Cost-sensitive learning;
D O I
暂无
中图分类号
学科分类号
摘要
Classifying non-stationary and imbalanced data streams encompasses two important challenges, namely concept drift and class imbalance. Concept drift is changes in the underlying function being learnt, and class imbalance is vast difference between the numbers of instances in different classes of data. Class imbalance is an obstacle for the efficiency of most classifiers. Previous methods for classifying non-stationary and imbalanced data streams mainly focus on batch solutions, in which the classification model is trained using a chunk of data. Here, we propose two online classifiers. The classifiers are one-layer NNs. In the proposed classifiers, class imbalance is handled with two separate cost-sensitive strategies. The first one incorporates a fixed and the second one an adaptive misclassification cost matrix. The proposed classifiers are evaluated on 3 synthetic and 8 real-world datasets. The results show statistically significant improvements in imbalanced data metrics.
引用
收藏
页码:1283 / 1295
页数:12
相关论文
共 37 条
[1]  
Sun J(2011)Dynamic financial distress prediction using instance selection for the disposal of concept drift Expert Syst Appl 38 2566-2576
[2]  
Li H(2011)A robust incremental learning method for non-stationary environments Neurocomputing 74 1800-1808
[3]  
Martínez-Rego D(2011)An adaptive classifier for data streams Pattern Recognit 44 78-96
[4]  
Pérez-Sánchez B(2011)Incremental learning of concept drift in nonstationary environments IEEE Trans Neural Netw 22 1517-1531
[5]  
Fontenla-Romero O(2011)Classification using streaming random forests IEEE Trans Knowl Data Eng 23 22-36
[6]  
Alonso-Betanzos A(2011)Classification and novel class detection in concept-drifting data streams under time constraints IEEE Trans Knowl Data Eng 23 859-874
[7]  
Pavlidis NG(2010)Towards incremental learning of nonstationary imbalanced data stream: a multiple selectively recursive approach Evol Syst 2 35-50
[8]  
Tasoulis DK(2009)Learning from imbalanced data IEEE Trans Knowl Data Eng 21 1263-1284
[9]  
Adams NM(2009)A joint investigation of misclassification treatments and imbalanced datasets on neural network performance Neural Comput Appl 18 689-706
[10]  
Hand DJ(2007)A data reduction approach for resolving the imbalanced data issue in functional genomics Neural Comput Appl 16 295-306