BAFusion: Bidirectional Attention Fusion for 3D Object Detection Based on LiDAR and Camera

被引:4
|
作者
Liu, Min [1 ]
Jia, Yuanjun [2 ]
Lyu, Youhao [1 ]
Dong, Qi [2 ]
Yang, Yanyu [2 ]
机构
[1] Univ Sci & Technol China, Inst Adv Technol, Hefei 230088, Peoples R China
[2] China Acad Elect & Informat Technol, Beijing 100041, Peoples R China
关键词
3D object detection; LiDAR-camera fusion; cross attention;
D O I
10.3390/s24144718
中图分类号
O65 [分析化学];
学科分类号
070302 ; 081704 ;
摘要
3D object detection is a challenging and promising task for autonomous driving and robotics, benefiting significantly from multi-sensor fusion, such as LiDAR and cameras. Conventional methods for sensor fusion rely on a projection matrix to align the features from LiDAR and cameras. However, these methods often suffer from inadequate flexibility and robustness, leading to lower alignment accuracy under complex environmental conditions. Addressing these challenges, in this paper, we propose a novel Bidirectional Attention Fusion module, named BAFusion, which effectively fuses the information from LiDAR and cameras using cross-attention. Unlike the conventional methods, our BAFusion module can adaptively learn the cross-modal attention weights, making the approach more flexible and robust. Moreover, drawing inspiration from advanced attention optimization techniques in 2D vision, we developed the Cross Focused Linear Attention Fusion Layer (CFLAF Layer) and integrated it into our BAFusion pipeline. This layer optimizes the computational complexity of attention mechanisms and facilitates advanced interactions between image and point cloud data, showcasing a novel approach to addressing the challenges of cross-modal attention calculations. We evaluated our method on the KITTI dataset using various baseline networks, such as PointPillars, SECOND, and Part-A2, and demonstrated consistent improvements in 3D object detection performance over these baselines, especially for smaller objects like cyclists and pedestrians. Our approach achieves competitive results on the KITTI benchmark.
引用
收藏
页数:26
相关论文
共 50 条
  • [41] Camera and LiDAR analysis for 3D object detection in foggy weather conditions
    Nguyen Anh Minh Mai
    Duthon, Pierre
    Salmane, Pascal Housam
    Khoudour, Louahdi
    Crouzil, Alain
    Velastin, Sergio A.
    2022 12TH INTERNATIONAL CONFERENCE ON PATTERN RECOGNITION SYSTEMS (ICPRS), 2022,
  • [42] CLFusion:3D Semantic Segmentation Based on Camera and Lidar Fusion
    Wang, Tianyue
    Song, Rujun
    Xiao, Zhuoling
    Yan, Bo
    Qin, Haojie
    He, Di
    2024 IEEE INTERNATIONAL SYMPOSIUM ON CIRCUITS AND SYSTEMS, ISCAS 2024, 2024,
  • [43] TEMPORAL AXIAL ATTENTION FOR LIDAR-BASED 3D OBJECT DETECTION IN AUTONOMOUS DRIVING
    Carranza-Garcia, Manuel
    Riquelme, Jose C.
    Zakhor, Avideh
    2022 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING, ICIP, 2022, : 201 - 205
  • [44] Center-Aware 3D Object Detection with Attention Mechanism Based on Roadside LiDAR
    Shi, Haobo
    Hou, Dezao
    Li, Xiyao
    SUSTAINABILITY, 2023, 15 (03)
  • [45] 3D object detection based on fusion of point cloud and image by mutual attention
    Chen J.-Y.
    Bai T.-Y.
    Zhao L.
    Guangxue Jingmi Gongcheng/Optics and Precision Engineering, 2021, 29 (09): : 2247 - 2254
  • [46] 3D Object Detection Based on Attention and Multi-Scale Feature Fusion
    Liu, Minghui
    Ma, Jinming
    Zheng, Qiuping
    Liu, Yuchen
    Shi, Gang
    SENSORS, 2022, 22 (10)
  • [47] Camera and LiDAR Fusion for Robust 3D Person Detection in Indoor Environments
    Silva, Carlos A.
    Dogru, Sedat
    Marques, Lino
    2023 IEEE INTERNATIONAL CONFERENCE ON AUTONOMOUS ROBOT SYSTEMS AND COMPETITIONS, ICARSC, 2023, : 187 - 192
  • [48] 3D LiDAR and Color Camera Data Fusion
    Ding, Yuqi
    Liu, Jiaming
    Ye, Jinwei
    Xiang, Weidong
    Wu, Hsiao-Chun
    Busch, Costas
    2020 IEEE INTERNATIONAL SYMPOSIUM ON BROADBAND MULTIMEDIA SYSTEMS AND BROADCASTING (BMSB), 2020,
  • [49] 3D Object Detection and Tracking Based on Lidar-Camera Fusion and IMM-UKF Algorithm Towards Highway Driving
    Nie, Chang
    Ju, Zhiyang
    Sun, Zhifeng
    Zhang, Hui
    IEEE TRANSACTIONS ON EMERGING TOPICS IN COMPUTATIONAL INTELLIGENCE, 2023, 7 (04): : 1242 - 1252
  • [50] FusionPainting: Multimodal Fusion with Adaptive Attention for 3D Object Detection
    Xu, Shaoqing
    Zhou, Dingfu
    Fang, Jin
    Yin, Junbo
    Bin, Zhou
    Zhang, Liangjun
    2021 IEEE INTELLIGENT TRANSPORTATION SYSTEMS CONFERENCE (ITSC), 2021, : 3047 - 3054