CrossEI: Boosting Motion-Oriented Object Tracking With an Event Camera

被引：0

作者：

Chen, Zhiwen ^{[1
]}

Wu, Jinjian ^{[1
]}

Dong, Weisheng ^{[1
]}

Li, Leida ^{[1
]}

Shi, Guangming ^{[1
]}

机构：

[1] Xidian Univ, Sch Artificial Intelligence, Xian 710071, Peoples R China

来源：

IEEE TRANSACTIONS ON IMAGE PROCESSING | 2025年 / 34卷

关键词：

Tracking; Cameras; Feature extraction; Semantics; Object tracking; Motion estimation; Modulation; Dynamics; Sensitivity; Benchmark testing; Event camera; event-image fusion; object tracking; VISION;

D O I：

10.1109/TIP.2024.3505672

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

With the differential sensitivity and high time resolution, event cameras can record detailed motion clues, which form a complementary advantage with frame-based cameras to enhance the object tracking, especially in challenging dynamic scenes. However, how to better match heterogeneous event-image data and exploit rich complementary cues from them still remains an open issue. In this paper, we align event-image modalities by proposing a motion adaptive event sampling method, and we revisit the cross-complementarities of event-image data to design a bidirectional-enhanced fusion framework. Specifically, this sampling strategy can adapt to different dynamic scenes and integrate aligned event-image pairs. Besides, we design an image-guided motion estimation unit for extracting explicit instance-level motions, aiming at refining the uncertain event clues to distinguish primary objects and background. Then, a semantic modulation module is devised to utilize the enhanced object motion to modify the image features. Coupled with these two modules, this framework learns both the high motion sensitivity of events and the full texture of images to achieve more accurate and robust tracking. The proposed method is easily embedded in existing tracking pipelines, and trained end-to-end. We evaluate it on four large benchmarks, i.e. FE108, VisEvent, FE240hz and CoeSot. Extensive experiments demonstrate our method achieves state-of-the-art performance, and large improvements are pointed as contributions by our sampling strategy and fusion concept.

引用

页码：73 / 84

页数：12

共 63 条

[61] Zhu AZ, 2018, ROBOTICS: SCIENCE AND SYSTEMS XIV
[62] Unsupervised Event-based Learning of Optical Flow, Depth, and Egomotion
Zhu, Alex Zihao
Yuan, Liangzhe
Chaney, Kenneth
Daniilidis, Kostas
[J]. 2019 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2019), 2019, : 989 - 997
[63] Zhu Z, 2022, Advances in neural information processing systems (NeurIPS)

← 1 2 3 4 5 6 7 →