PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 19, 2026World Electric Vehicle Journal0 citationsOpen Access

NMLoNet: An End-to-End Intelligent Vehicle Localization Network Using Navigation Maps

View Full Paper
QYQingtong YuanYLYicheng Li

Key Points

  • The research aims to enhance vehicle localization accuracy using navigation maps, addressing the challenges posed by traditional mapping methods.
  • Developed NMLoNet architecture for vehicle localization
  • Introduced a Deformable Attention Module to enhance BEV feature extraction
  • Incorporated vector map constraints to align BEV and navigation map features
  • Designed a multi-level cross-modal feature registration mechanism
  • Achieved an 11% improvement in localization accuracy under monocular settings
  • Improved accuracy by 24% with surround-view configurations
  • Demonstrated robustness in complex and dynamic driving conditions

Abstract

Accurate and reliable localization is crucial for advanced autonomous driving systems. Traditional high-precision localization approaches rely on meticulously annotated high-definition (HD) maps and employ visual-geometric methods to derive accurate pose information. However, the construction, maintenance, and updating of HD maps are costly and time-consuming. In contrast, localization using publicly available navigation maps provides a low-cost and scalable alternative. Existing methods typically align BEV (Bird’s-Eye-View) features extracted from surround-view images with navigation maps to obtain localization results. Although such approaches can achieve high accuracy, they often neglect the inherent modality gap between BEV features and navigation maps, leading to localization errors. To address this issue, we propose NMLoNet: An End-to-End Intelligent Vehicle Localization Network Using Navigation Maps. The proposed method exploits road semantic elements to effectively bridge the modality gap between BEV representations and navigation maps. Specifically, a Deformable Attention Module is introduced after BEV feature extraction to capture long-range dependencies among BEV features. Furthermore, we innovatively incorporate vector map constraints to minimize the discrepancy between BEV and navigation map features. In addition, a multi-level cross-modal feature registration mechanism is designed to achieve more precise alignment between BEV and map representations. Extensive experiments on the nuScenes and Argoverse datasets demonstrate that NMLoNet achieves state-of-the-art performance, improving localization accuracy by approximately 11% under monocular settings and 24% under surround-view configurations. Moreover, the proposed network maintains robust localization performance in complex and highly dynamic driving environments.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Yuan et al. (2026) studied this question.

synapsesocial.com/papers/69bb92be496e729e62980413https://doi.org/10.3390/wevj17030150
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1PoseNet: A Convolutional Network for Real-Time 6-DOF Camera Relocalization2015 · 2,393 citations
  2. 2Visual-lidar odometry and mapping: low-drift, robust, and fast2015 · 758 citations
  3. 3Pixel-Perfect Structure-from-Motion with Featuremetric Refinement2021 · 165 citations
  4. 4A Review of Motion Planning for Highway Autonomous Driving2019 · 675 citations
  5. 5Untitled21 citations