A Context-Aware Loss Function for Action Spotting in Soccer Videos

被引:46
作者
Cioppa, Anthony [1 ]
Deliege, Adrien [1 ]
Giancola, Silvio [2 ]
Ghanem, Bernard [2 ]
Van Droogenbroeck, Marc [1 ]
Gade, Rikke [3 ]
Moeslund, Thomas B. [3 ]
机构
[1] Univ Liege, Liege, Belgium
[2] KAUST, Thuwal, Saudi Arabia
[3] Aalborg Univ, Aalborg, Denmark
来源
2020 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2020) | 2020年
关键词
D O I
10.1109/CVPR42600.2020.01314
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
In video understanding, action spotting consists in temporally localizing human-induced events annotated with single timestamps. In this paper, we propose a novel loss function that specifically considers the temporal context naturally present around each action, rather than focusing on the single annotated frame to spot. We benchmark our loss on a large dataset of soccer videos, SoccerNet, and achieve an improvement of 12.8% over the baseline. We show the generalization capability of our loss for generic activity proposals and detection on ActivityNet, by spotting the beginning and the end of each activity. Furthermore, we provide an extended ablation study and display challenging cases for action spotting in soccer videos. Finally, we qualitatively illustrate how our loss induces a precise temporal understanding of actions and show how such semantic knowledge can be used for automatic highlights generation.
引用
收藏
页码:13123 / 13133
页数:11
相关论文
共 69 条
[1]   Action Search: Spotting Actions in Videos and Its Application to Temporal Action Localization [J].
Alwassel, Humam ;
Heilbron, Fabian Caba ;
Ghanem, Bernard .
COMPUTER VISION - ECCV 2018, PT IX, 2018, 11213 :253-269
[2]  
Alwassel Humam, EUR C COMP VIS ECCV
[3]  
[Anonymous], 2018, IEEE C COMP VIS PATT
[4]  
[Anonymous], 2017, IEEE C COMP VIS PATT
[5]  
[Anonymous], 2015, IEEE C COMP VIS PATT
[6]  
Baccouche M., 2010, INT C ART NEUR NETW
[7]  
Bettadapura Vinay, 2016, ACM INT C MULT ACM M
[8]  
Bridgeman Lewis, 2019, IEEE C COMP VIS PATT
[9]   SST: Single-Stream Temporal Action Proposals [J].
Buch, Shyamal ;
Escorcia, Victor ;
Shen, Chuanqi ;
Ghanem, Bernard ;
Niebles, Juan Carlos .
30TH IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2017), 2017, :6373-6382
[10]  
Buch Shyamal, 2017, BRIT MACH VIS C BMVC