BAFusion: Bidirectional Attention Fusion for 3D Object Detection Based on LiDAR and Camera

被引:1
|
作者
Liu, Min [1 ]
Jia, Yuanjun [2 ]
Lyu, Youhao [1 ]
Dong, Qi [2 ]
Yang, Yanyu [2 ]
机构
[1] Univ Sci & Technol China, Inst Adv Technol, Hefei 230088, Peoples R China
[2] China Acad Elect & Informat Technol, Beijing 100041, Peoples R China
关键词
3D object detection; LiDAR-camera fusion; cross attention;
D O I
10.3390/s24144718
中图分类号
O65 [分析化学];
学科分类号
070302 ; 081704 ;
摘要
3D object detection is a challenging and promising task for autonomous driving and robotics, benefiting significantly from multi-sensor fusion, such as LiDAR and cameras. Conventional methods for sensor fusion rely on a projection matrix to align the features from LiDAR and cameras. However, these methods often suffer from inadequate flexibility and robustness, leading to lower alignment accuracy under complex environmental conditions. Addressing these challenges, in this paper, we propose a novel Bidirectional Attention Fusion module, named BAFusion, which effectively fuses the information from LiDAR and cameras using cross-attention. Unlike the conventional methods, our BAFusion module can adaptively learn the cross-modal attention weights, making the approach more flexible and robust. Moreover, drawing inspiration from advanced attention optimization techniques in 2D vision, we developed the Cross Focused Linear Attention Fusion Layer (CFLAF Layer) and integrated it into our BAFusion pipeline. This layer optimizes the computational complexity of attention mechanisms and facilitates advanced interactions between image and point cloud data, showcasing a novel approach to addressing the challenges of cross-modal attention calculations. We evaluated our method on the KITTI dataset using various baseline networks, such as PointPillars, SECOND, and Part-A2, and demonstrated consistent improvements in 3D object detection performance over these baselines, especially for smaller objects like cyclists and pedestrians. Our approach achieves competitive results on the KITTI benchmark.
引用
收藏
页数:26
相关论文
共 50 条
  • [1] A LiDAR-Camera Fusion 3D Object Detection Algorithm
    Liu, Leyuan
    He, Jian
    Ren, Keyan
    Xiao, Zhonghua
    Hou, Yibin
    [J]. INFORMATION, 2022, 13 (04)
  • [2] CLOCs: Camera-LiDAR Object Candidates Fusion for 3D Object Detection
    Pang, Su
    Morris, Daniel
    Radha, Hayder
    [J]. 2020 IEEE/RSJ INTERNATIONAL CONFERENCE ON INTELLIGENT ROBOTS AND SYSTEMS (IROS), 2020, : 10386 - 10393
  • [3] 3D Vehicle Detection Based on LiDAR and Camera Fusion
    Cai, Yingfeng
    Zhang, Tiantian
    Wang, Hai
    Li, Yicheng
    Liu, Qingchao
    Chen, Xiaobo
    [J]. AUTOMOTIVE INNOVATION, 2019, 2 (04) : 276 - 283
  • [4] 3D Vehicle Detection Based on LiDAR and Camera Fusion
    Yingfeng Cai
    Tiantian Zhang
    Hai Wang
    Yicheng Li
    Qingchao Liu
    Xiaobo Chen
    [J]. Automotive Innovation, 2019, 2 : 276 - 283
  • [5] SupFusion: Supervised LiDAR-Camera Fusion for 3D Object Detection
    Qin, Yiran
    Wang, Chaoqun
    Kang, Zijian
    Ma, Ningning
    Li, Zhen
    Zhang, Ruimao
    [J]. 2023 IEEE/CVF INTERNATIONAL CONFERENCE ON COMPUTER VISION (ICCV 2023), 2023, : 21957 - 21967
  • [6] BEV Space 3D Object Detection Algorithm Based on Fusion of Infrared Camera and LiDAR
    Wang Wuyue
    Xu Zhaofei
    Qu Chunyan
    Lin Ying
    Chen Yufeng
    Liao Jian
    [J]. ACTA PHOTONICA SINICA, 2024, 53 (01)
  • [7] DNN Based Camera and Lidar Fusion Framework for 3D Object Recognition
    Zhang, K.
    Wang, S. J.
    Ji, L.
    Wang, C.
    [J]. 2020 4TH INTERNATIONAL CONFERENCE ON MACHINE VISION AND INFORMATION TECHNOLOGY (CMVIT 2020), 2020, 1518
  • [8] A Frustum-based probabilistic framework for 3D object detection by fusion of LiDAR and camera data
    Gong, Zheng
    Lin, Haojia
    Zhang, Dedong
    Luo, Zhipeng
    Zelek, John
    Chen, Yiping
    Nurunnabi, Abdul
    Wang, Cheng
    Li, Jonathan
    [J]. ISPRS JOURNAL OF PHOTOGRAMMETRY AND REMOTE SENSING, 2020, 159 : 90 - 100
  • [9] TransFusion: Robust LiDAR-Camera Fusion for 3D Object Detection with Transformers
    Bai, Xuyang
    Hu, Zeyu
    Zhu, Xinge
    Huang, Qingqiu
    Chen, Yilun
    Fu, Hangbo
    Tai, Chiew-Lan
    [J]. 2022 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2022), 2022, : 1080 - 1089
  • [10] Fusion of 3D LIDAR and Camera Data for Object Detection in Autonomous Vehicle Applications
    Zhao, Xiangmo
    Sun, Pengpeng
    Xu, Zhigang
    Min, Haigen
    Yu, Hongkai
    [J]. IEEE SENSORS JOURNAL, 2020, 20 (09) : 4901 - 4913