Scene categorization via contextual visual words

被引:68
作者
Qin, Jianzhao [1 ]
Yung, Nelson H. C. [1 ]
机构
[1] Univ Hong Kong, Dept Elect & Elect Engn, Lab Intelligent Transportat Syst Res, Hong Kong, Hong Kong, Peoples R China
关键词
Scene categorization; Contextual visual words; Context based vision; Pattern recognition; CLASSIFICATION; OBJECTS; IMAGES;
D O I
10.1016/j.patcog.2009.11.009
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
In this paper, we propose a novel scene categorization method based on contextual visual words In the proposed method, we extend the traditional 'bags of visual words' model by introducing contextual information from the coarser scale and neighborhood regions to the local region of interest based on unsupervised learning. The introduced contextual information provides useful information or cue about the region of interest, which can reduce the ambiguity when employing visual words to represent the local regions The improved visual words representation of the scene image is capable of enhancing the categorization performance The proposed method is evaluated over three scene classification datasets, with 8, 13 and 15 scene categories, respectively, using 10-fold cross-validation The experimental results show that the proposed method achieves 90 30%, 87 63% and 85.16% recognition success for Dataset 1,2 and 3, respectively, which significantly outperforms the methods based on the visual words that only represent the local information in the statistical manner We also compared the proposed method with three representative scene categorization methods The result confirms the superiority of the proposed method. (C) 2009 Elsevier Ltd All rights reserved.
引用
收藏
页码:1874 / 1888
页数:15
相关论文
共 40 条
[11]  
Fergus R, 2005, IEEE I CONF COMP VIS, P1816
[12]  
Grauman K, 2005, IEEE I CONF COMP VIS, P1458
[13]  
Harris C, 1988, ALVEY VISION C, V15, P10, DOI DOI 10.5244/C.2.23
[14]  
He XM, 2004, PROC CVPR IEEE, P695
[15]  
Heitz G., 2008, LEARNING SPATIAL CON
[16]   Putting objects in perspective [J].
Hoiem, Derek ;
Efros, Alexei A. ;
Hebert, Martial .
INTERNATIONAL JOURNAL OF COMPUTER VISION, 2008, 80 (01) :3-15
[17]  
JIAYAN J, 2004, P 2004 IEEE COMP SOC
[18]   Saliency, scale and image description [J].
Kadir, T ;
Brady, M .
INTERNATIONAL JOURNAL OF COMPUTER VISION, 2001, 45 (02) :83-105
[19]  
Kumar S, 2005, IEEE I CONF COMP VIS, P1284
[20]   Discriminative random fields: A discriminative framework for contextual interaction in classification [J].
Kumar, S ;
Hebert, M .
NINTH IEEE INTERNATIONAL CONFERENCE ON COMPUTER VISION, VOLS I AND II, PROCEEDINGS, 2003, :1150-1157