Interaction-Aware Spatio-Temporal Pyramid Attention Networks for Action Classification

被引:76
作者
Du, Yang [1 ,2 ,3 ]
Yuan, Chunfeng [2 ]
Li, Bing [2 ]
Zhao, Lili [3 ]
Li, Yangxi [4 ]
Hu, Weiming [2 ]
机构
[1] Univ Chinese Acad Sci, Beijing, Peoples R China
[2] Chinese Acad Sci, Inst Automat, Natl Lab Pattern Recognit, CAS Ctr Excellence Brain Sci & Intelligence Techn, Beijing, Peoples R China
[3] Meitu, Mainland, Peoples R China
[4] Natl Comp Network Emergency Response Tech Team Co, Ho Chi Minh City, Vietnam
来源
COMPUTER VISION - ECCV 2018, PT XVI | 2018年 / 11220卷
关键词
D O I
10.1007/978-3-030-01270-0_23
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Local features at neighboring spatial positions in feature maps have high correlation since their receptive fields are often overlapped. Self-attention usually uses the weighted sum (or other functions) with internal elements of each local feature to obtain its weight score, which ignores interactions among local features. To address this, we propose an effective interaction-aware self-attention model inspired by PCA to learn attention maps. Furthermore, since different layers in a deep network capture feature maps of different scales, we use these feature maps to construct a spatial pyramid and then utilize multi-scale information to obtain more accurate attention scores, which are used to weight the local features in all spatial positions of feature maps to calculate attention maps. Moreover, our spatial pyramid attention is unrestricted to the number of its input feature maps so it is easily extended to a spatiotemporal version. Finally, our model is embedded in general CNNs to form end-to-end attention networks for action classification. Experimental results show that our method achieves the state-of-the-art results on the UCF101, HMDB51 and untrimmed Charades.
引用
收藏
页码:388 / 404
页数:17
相关论文
共 58 条
[1]  
Abadi M., 2015, TensorFlow: Large-scale machine learning on heterogeneous systems.
[2]  
Abu-El-Haija S., 2016, CORR
[3]  
[Anonymous], 2016, ECCV
[4]  
[Anonymous], 2017, 31 AAAI C ART INT AA
[5]  
[Anonymous], 2014, arXiv
[6]  
[Anonymous], 2015, CVPR
[7]  
[Anonymous], 2015, P IEEE INT C COMPUTE
[8]  
[Anonymous], 2014, NIPS
[9]  
[Anonymous], 2015, CORR
[10]  
[Anonymous], 1997, Neural Computation