Tracking by Instance Detection: A Meta-Learning Approach

被引：157

作者：

Wang, Guangting ^{[1
,3
]}

Luo, Chong ^{[2
]}

Sun, Xiaoyan ^{[2
]}

Xiong, Zhiwei ^{[1
]}

Zeng, Wenjun ^{[2
]}

机构：

[1] Univ Sci & Thchnol China, Hefei, Anhui, Peoples R China

[2] Microsoft Res Asia, Beijing, Peoples R China

[3] MSRA, Beijing, Peoples R China

来源：

2020 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR) | 2020年

关键词：

D O I：

10.1109/CVPR42600.2020.00632

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

We consider the tracking problem as a special type of object detection problem, which we call instance detection. With proper initialization, a detector can be quickly converted into a tracker by learning the new instance from a single image. We find that model-agnostic meta-learning (MAML) offers a strategy to initialize the detector that satisfies our needs. We propose a principled three-step approach to build a high-performance tracker. First, pick any modern object detector trained with gradient descent. Second, conduct offline training (or initialization) with MAML. Third, perform domain adaptation using the initial frame. We follow this procedure to build two trackers, named RetinaMAML and FCOS-MAML, based on two modern detectors RetinaNet and FCOS. Evaluations on four benchmarks show that both trackers are competitive against state-ofthe-art trackers. On OTB-100, Retina-MAML achieves the highest ever AUC of 0.712. On TrackingNet, FCOS-MAML ranks the first on the leader board with an AUC of 0.757 and the normalized precision of 0.822. Both trackers run in real-time at 40 FPS.

引用

页码：6287 / 6296

页数：10

共 42 条

[1]

[Anonymous], IEEE T PATTERN ANAL

[2]

[Anonymous], 2015, ACS SYM SER

[3]

[Anonymous], 2017, META SGD LEARNING LE

[4]

Antoniou A, 2018, TRAIN YOUR MAML

[5] Fully-Convolutional Siamese Networks for Object Tracking [J].

Bertinetto, Luca ;

Valmadre, Jack ;

Henriques, Joao F. ;

Vedaldi, Andrea ;

Torr, Philip H. S. .

COMPUTER VISION - ECCV 2016 WORKSHOPS, PT II, 2016, 9914 :850-865

[6] Learning Discriminative Model Prediction for Tracking [J].

Bhat, Goutam ;

Danelljan, Martin ;

Van Gool, Luc ;

Timofte, Radu .

2019 IEEE/CVF INTERNATIONAL CONFERENCE ON COMPUTER VISION (ICCV 2019), 2019, :6181-6190

[7] Unveiling the Power of Deep Tracking [J].

Bhat, Goutam ;

Johnander, Joakim ;

Danelljan, Martin ;

Khan, Fahad Shahbaz ;

Felsberg, Michael .

COMPUTER VISION - ECCV 2018, PT II, 2018, 11206 :493-509

[8] Deep Meta Learning for Real-Time Target-Aware Visual Tracking [J].

Choi, Janghoon ;

Kwon, Junseok ;

Lee, Kyoung Mu .

2019 IEEE/CVF INTERNATIONAL CONFERENCE ON COMPUTER VISION (ICCV 2019), 2019, :911-920

[9] ATOM: Accurate Tracking by Overlap Maximization [J].

Danelljan, Martin ;

Bhat, Goutam ;

Khan, Fahad Shahbaz ;

Felsberg, Michael .

2019 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2019), 2019, :4655-4664

[10] ECO: Efficient Convolution Operators for Tracking [J].

Danelljan, Martin ;

Bhat, Goutam ;

Khan, Fahad Shahbaz ;

Felsberg, Michael .

30TH IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2017), 2017, :6931-6939

← 1 2 3 4 5 →