Swin transformer based vehicle detection in undisciplined traffic environment

被引:33
作者
Deshmukh, Prashant [1 ]
Satyanarayana, G. S. R. [1 ,2 ]
Majhi, Sudhan [3 ]
Sahoo, Upendra Kumar [1 ]
Das, Santos Kumar [1 ]
机构
[1] Natl Inst Technol Rourkela, Dept Elect & Commun Engn, Rourkela, India
[2] Vignans Fdn Sci Technol & Res, Dept Elect & Commun Engn, Guntur, India
[3] Indian Inst Sci, Dept Elect Commun Engn, Bangalore, India
关键词
Deep learning; Undisciplined traffic environment; Visual transformer; Vehicle detection; DETECTION SYSTEM; OBJECT DETECTION; CLASSIFICATION;
D O I
10.1016/j.eswa.2022.118992
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Intelligent vehicle detection (IVD) plays a prominent role in evolving an intelligent traffic management system (ITMS). It can help to decrease the average waiting time at the traffic post, save fuel consumption, control traffic congestion, decrease accident rates, and build up human safety. Recent developments in the artificial intelligence (AI) domain have increased the demand for IVD in the undisciplined traffic environment, which is a usual condition in developing countries. IVD is a difficult task in an undisciplined traffic environment because different vehicle categories travel very close to each other on the roads and do not follow traffic rules. Previously, several convolutional neural network (CNN) based deep learning (DL), and visual transformer-based techniques for vehicle and object detection have been presented. They are complex and do not accurately extract multi-scale features due to the involvement of existing CNN feature extraction backbones. Also, most techniques failed to account for an undisciplined traffic environment due to the unavailability of labeled vehicle datasets. Therefore, this paper proposes a swin transformer-based vehicle detection (STVD) framework in an undisciplined traffic environment. Swin transformer (ST) wholly exchanges information within and between image patches and provides hierarchical feature maps, effectively alleviating the multi-scale feature extraction problem. A bi-directional feature pyramid network (BIFPN) is presented, which combines low -resolution features with high-resolution features in a bidirectional way and provides robust multi-scale features with different scales and resolutions. A fully connected vehicle detection head (FCVDH) is applied to improve the matching relationship between vehicle sizes and the BIFPN hierarchy. FCVDH predicts the locations and categories of vehicles in the input image. STVD is analyzed, experimented, and measured over realistic traffic data. Also, it is compared with the existing state-of-the-art vehicle detection methods. It achieves 91.32% detection accuracy on diverse traffic labeled dataset (DTLD), 87.4% on IITM-hetra, and 88.45% on KITTI datasets.
引用
收藏
页数:13
相关论文
共 69 条
[21]  
Han G., 2022, IEEE C COMPUTER VISI, P5321
[22]   A Survey on Vision Transformer [J].
Han, Kai ;
Wang, Yunhe ;
Chen, Hanting ;
Chen, Xinghao ;
Guo, Jianyuan ;
Liu, Zhenhua ;
Tang, Yehui ;
Xiao, An ;
Xu, Chunjing ;
Xu, Yixing ;
Yang, Zhaohui ;
Zhang, Yiman ;
Tao, Dacheng .
IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2023, 45 (01) :87-110
[23]   A VEHICLE DETECTION SYSTEM BASED ON HAAR AND TRIANGLE FEATURES [J].
Haselhoff, Anselm ;
Kummert, Anton .
2009 IEEE INTELLIGENT VEHICLES SYMPOSIUM, VOLS 1 AND 2, 2009, :261-266
[24]  
He KM, 2020, IEEE T PATTERN ANAL, V42, P386, DOI [10.1109/ICCV.2017.322, 10.1109/TPAMI.2018.2844175]
[25]   Deep Residual Learning for Image Recognition [J].
He, Kaiming ;
Zhang, Xiangyu ;
Ren, Shaoqing ;
Sun, Jian .
2016 IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2016, :770-778
[26]   A new method of moving object detection using adaptive filter [J].
Hsia, Chih-Hsien ;
Wu, Tsung-Cheng ;
Chiang, Jen-Shiun .
JOURNAL OF REAL-TIME IMAGE PROCESSING, 2017, 13 (02) :311-325
[27]   Symmetrical SURF and Its Applications to Vehicle Detection and Vehicle Make and Model Recognition [J].
Hsieh, Jun-Wei ;
Chen, Li-Chih ;
Chen, Duan-Yu .
IEEE TRANSACTIONS ON INTELLIGENT TRANSPORTATION SYSTEMS, 2014, 15 (01) :6-20
[28]  
Hsu SC, 2018, PROC INT WORKSH ADV
[29]   SINet: A Scale-Insensitive Convolutional Neural Network for Fast Vehicle Detection [J].
Hu, Xiaowei ;
Xu, Xuemiao ;
Xiao, Yongjie ;
Chen, Hao ;
He, Shengfeng ;
Qin, Jing ;
Heng, Pheng-Ann .
IEEE TRANSACTIONS ON INTELLIGENT TRANSPORTATION SYSTEMS, 2019, 20 (03) :1010-1019
[30]   Densely Connected Convolutional Networks [J].
Huang, Gao ;
Liu, Zhuang ;
van der Maaten, Laurens ;
Weinberger, Kilian Q. .
30TH IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2017), 2017, :2261-2269