PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
July 20, 2024Sensors13 citationsOpen Access

BAFusion: Bidirectional Attention Fusion for 3D Object Detection Based on LiDAR and Camera

View Full Paper
LMLee Jun MinUniversity of Science and Technology of ChinaYJYuanjun JiaChina Academy of Information and Communications TechnologyYLYunmiao LyuUniversity of Science and Technology of China

Key Points

Key points are not available for this paper at this time.

Abstract

3D object detection is a challenging and promising task for autonomous driving and robotics, benefiting significantly from multi-sensor fusion, such as LiDAR and cameras. Conventional methods for sensor fusion rely on a projection matrix to align the features from LiDAR and cameras. However, these methods often suffer from inadequate flexibility and robustness, leading to lower alignment accuracy under complex environmental conditions. Addressing these challenges, in this paper, we propose a novel Bidirectional Attention Fusion module, named BAFusion, which effectively fuses the information from LiDAR and cameras using cross-attention. Unlike the conventional methods, our BAFusion module can adaptively learn the cross-modal attention weights, making the approach more flexible and robust. Moreover, drawing inspiration from advanced attention optimization techniques in 2D vision, we developed the Cross Focused Linear Attention Fusion Layer (CFLAF Layer) and integrated it into our BAFusion pipeline. This layer optimizes the computational complexity of attention mechanisms and facilitates advanced interactions between image and point cloud data, showcasing a novel approach to addressing the challenges of cross-modal attention calculations. We evaluated our method on the KITTI dataset using various baseline networks, such as PointPillars, SECOND, and Part-A

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Min et al. (2024) studied this question.

synapsesocial.com/papers/68e5fa56b6db64358758e192https://doi.org/10.3390/s24144718
Ask AI
Helpful
Bookmark
Share
View Full Paper