Online cost-sensitive neural network classifiers for non-stationary and imbalanced data streams

被引:30
作者
Ghazikhani, Adel [1 ]
Monsefi, Reza [1 ]
Yazdi, Hadi Sadoghi [1 ]
机构
[1] Ferdowsi Univ Mashhad, Dept Comp Engn, Mashhad, Iran
关键词
Data stream classification; One-layer neural network; Concept drift; Imbalanced data; Online classification; Cost-sensitive learning; PREDICTION;
D O I
10.1007/s00521-012-1071-6
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Classifying non-stationary and imbalanced data streams encompasses two important challenges, namely concept drift and class imbalance. Concept drift is changes in the underlying function being learnt, and class imbalance is vast difference between the numbers of instances in different classes of data. Class imbalance is an obstacle for the efficiency of most classifiers. Previous methods for classifying non-stationary and imbalanced data streams mainly focus on batch solutions, in which the classification model is trained using a chunk of data. Here, we propose two online classifiers. The classifiers are one-layer NNs. In the proposed classifiers, class imbalance is handled with two separate cost-sensitive strategies. The first one incorporates a fixed and the second one an adaptive misclassification cost matrix. The proposed classifiers are evaluated on 3 synthetic and 8 real-world datasets. The results show statistically significant improvements in imbalanced data metrics.
引用
收藏
页码:1283 / 1295
页数:13
相关论文
共 29 条
[11]   Learning from Imbalanced Data [J].
He, Haibo ;
Garcia, Edwardo A. .
IEEE TRANSACTIONS ON KNOWLEDGE AND DATA ENGINEERING, 2009, 21 (09) :1263-1284
[12]  
Harries M., 1999, SPLICE 2 COMP EVALUA
[13]  
Klinkenberg R., 2000, 17 INT C MACH LEARN
[14]   A joint investigation of misclassification treatments and imbalanced datasets on neural network performance [J].
Lan, Jyh-shyan ;
Berardi, Victor L. ;
Patuwo, B. Eddy ;
Hu, Michael .
NEURAL COMPUTING & APPLICATIONS, 2009, 18 (07) :689-706
[15]  
Lichtenwalter R, 2009, PAKDD
[16]  
Lichtenwalter R, 2009, PAKDD WORKSH DAT MIN
[17]  
Martínez-Rego D, 2011, NEUROCOMPUTING, V74, P1800, DOI [10.1016/j.neucom.2010.06.037, 10.1016/j.neucom.2010.06.03]
[18]   Classification and Novel Class Detection in Concept-Drifting Data Streams under Time Constraints [J].
Masud, Mohammad M. ;
Gao, Jing ;
Khan, Latifur ;
Han, Jiawei ;
Thuraisingham, Bhavani .
IEEE TRANSACTIONS ON KNOWLEDGE AND DATA ENGINEERING, 2011, 23 (06) :859-874
[19]  
Narasimhamurthy A, 2007, IASTED INT C ART INT
[20]  
NOAA, 2010, WEATH DAT