PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 12, 20240 citationsOpen Access

Eliminating Cross-modal Conflicts in BEV Space for LiDAR-Camera 3D Object Detection

View Full Paper
JFJiahui FuGCGao ChenZWZitian Wang

Key Points

Key points are not available for this paper at this time.

Abstract

Recent 3D object detectors typically utilize multi-sensor data and unify multi-modal features in the shared bird's-eye view (BEV) representation space. However, our empirical findings indicate that previous methods have limitations in generating fusion BEV features free from cross-modal conflicts. These conflicts encompass extrinsic conflicts caused by BEV feature construction and inherent conflicts stemming from heterogeneous sensor signals. Therefore, we propose a novel Eliminating Conflicts Fusion (ECFusion) method to explicitly eliminate the extrinsic/inherent conflicts in BEV space and produce improved multi-modal BEV features. Specifically, we devise a Semantic-guided Flow-based Alignment (SFA) module to resolve extrinsic conflicts via unifying spatial distribution in BEV space before fusion. Moreover, we design a Dissolved Query Recovering (DQR) mechanism to remedy inherent conflicts by preserving objectness clues that are lost in the fusion BEV feature. In general, our method maximizes the effective information utilization of each modality and leverages inter-modal complementarity. Our method achieves state-of-the-art performance in the highly competitive nuScenes 3D object detection dataset. The code is released at https://github.com/fjhzhixi/ECFusion.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Fu et al. (2024) studied this question.

synapsesocial.com/papers/68e747e0b6db6435876c0c0ahttps://doi.org/10.48550/arxiv.2403.07372
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Semantic-Enhanced and Temporally Refined Bidirectional BEV Fusion for LiDAR–Camera 3D Object Detection2025 · 2 citations
  2. 2CL-fusionBEV: 3D object detection method with camera-LiDAR fusion in Bird’s Eye View2024 · 16 citations
  3. 3GraphBEV: Towards Robust BEV Feature Alignment for Multi-Modal 3D Object Detection2024
  4. 4GPFusion: enhancing multi-modal 3D object detection via geometric pseudo-image fusion and adaptive feature fusion2026
  5. 5DepthFusion: Depth-Aware Hybrid Feature Fusion for LiDAR-Camera 3D Object Detection2025