Attention-Based Mean-Max Balance Assignment for Oriented Object Detection in Optical Remote Sensing Images

被引：0

作者：

Lin, Qifeng ^{[1
]}

Chen, Nuo ^{[1
]}

Huang, Haibin ^{[1
]}

Zhu, Daoye ^{[1
,2
]}

Fu, Gang ^{[3
]}

Chen, Chuanxi ^{[4
]}

Yu, Yuanlong ^{[1
]}

机构：

[1] Fuzhou Univ, Coll Comp & Data Sci, Fuzhou 350108, Peoples R China

[2] Univ Toronto, Dept Geog Geomatics & Environm, Mississauga, ON L5L 1C6, Canada

[3] Hong Kong Polytech Univ, Dept Comp, Hong Kong, Peoples R China

[4] Fujian Normal Univ, Coll Comp & Cyber Secur, Fuzhou 350117, Fujian, Peoples R China

来源：

IEEE TRANSACTIONS ON GEOSCIENCE AND REMOTE SENSING | 2025年 / 63卷

基金：

中国国家自然科学基金;

关键词：

Remote sensing; Detectors; Feature extraction; Object detection; Training; Semantics; Location awareness; Accuracy; Shape; Optical scattering; Attention feature fusion; label assignment; optical remote sensing (RS) images; oriented object detection;

D O I：

10.1109/TGRS.2025.3533553

中图分类号：

P3 [地球物理学]; P59 [地球化学];

学科分类号：

0708 ; 070902 ;

摘要：

For objects with arbitrary angles in optical remote sensing (RS) images, the oriented bounding box regression task often faces the problem of ambiguous boundaries between positive and negative samples. The statistical analysis of existing label assignment strategies reveals that anchors with low Intersection over Union (IoU) between ground truth (GT) may also accurately surround the GT after decoding. Therefore, this article proposes an attention-based mean-max balance assignment (AMMBA) strategy, which consists of two parts: mean-max balance assignment (MMBA) strategy and balance feature pyramid with attention (BFPA). MMBA employs the mean-max assignment (MMA) and balance assignment (BA) to dynamically calculate a positive threshold and adaptively match better positive samples for each GT for training. Meanwhile, to meet the need of MMBA for more accurate feature maps, we construct a BFPA module that integrates spatial and scale attention mechanisms to promote global information propagation. Combined with S2ANet, our AMMBA method can effectively achieve state-of-the-art performance, with a precision of 80.91% on the DOTA dataset in a simple plug-and-play fashion. Extensive experiments on three challenging optical RS image datasets (DOTA-v1.0, HRSC, and DIOR-R) further demonstrate the balance between precision and speed in single-stage object detectors. Our AMMBA has enough potential to assist all existing RS models in a simple way to achieve better detection performance. The code is available at https://github.com/promisekoloer/AMMBA.

引用

页数：15

共 59 条

[1] Anchor-Free Oriented Proposal Generator for Object Detection
Cheng, Gong
Wang, Jiabao
Li, Ke
Xie, Xingxing
Lang, Chunbo
Yao, Yanqing
Han, Junwei
[J]. IEEE TRANSACTIONS ON GEOSCIENCE AND REMOTE SENSING, 2022, 60
[2] Dynamic Head: Unifying Object Detection Heads with Attentions
Dai, Xiyang
Chen, Yinpeng
Xiao, Bin
Chen, Dongdong
Liu, Mengchen
Yuan, Lu
Zhang, Lei
[J]. 2021 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION, CVPR 2021, 2021, : 7369 - 7378
[3] Learning RoI Transformer for Oriented Object Detection in Aerial Images
Ding, Jian
Xue, Nan
Long, Yang
Xia, Gui-Song
Lu, Qikai
[J]. 2019 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2019), 2019, : 2844 - 2853
[4] OTA: Optimal Transport Assignment for Object Detection
Ge, Zheng
Liu, Songtao
Liu, Zeming
Yoshie, Osamu
Sun, Jian
[J]. 2021 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION, CVPR 2021, 2021, : 303 - 312
[5] Align Deep Features for Oriented Object Detection
Han, Jiaming
Ding, Jian
Li, Jie
Xia, Gui-Song
[J]. IEEE TRANSACTIONS ON GEOSCIENCE AND REMOTE SENSING, 2022, 60
[6] Deep Residual Learning for Image Recognition
He, Kaiming
Zhang, Xiangyu
Ren, Shaoqing
Sun, Jian
[J]. 2016 IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2016, : 770 - 778
[7] Hongkai Zhang, 2020, Computer Vision - ECCV 2020. 16th European Conference. Proceedings. Lecture Notes in Computer Science (LNCS 12360), P260, DOI 10.1007/978-3-030-58555-6_16
[8] Hou LP, 2022, AAAI CONF ARTIF INTE, P923
[9] Densely Connected Convolutional Networks
Huang, Gao
Liu, Zhuang
van der Maaten, Laurens
Weinberger, Kilian Q.
[J]. 30TH IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2017), 2017, : 2261 - 2269
[10] A General Gaussian Heatmap Label Assignment for Arbitrary-Oriented Object Detection
Huang, Zhanchao
Li, Wei
Xia, Xiang-Gen
Tao, Ran
[J]. IEEE TRANSACTIONS ON IMAGE PROCESSING, 2022, 31 : 1895 - 1910

← 1 2 3 4 5 6 →