PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 14, 2026Computer-Aided Civil and Infrastructure Engineering2 citationsOpen Access

Self-Supervised Adaptation of Dense Pavement Image Matching for Long-Term Road Monitoring under Challenging Visual Conditions

View Full Paper
CXChunliang XueKXKai XueMGMengxue Guo

Key Points

  • The aim is to enhance the matching of pavement images over time despite challenging visual conditions.
  • Utilized vehicle-mounted smartphones for image collection over five years and multiple regions.
  • Developed a self-supervised adaptation for dense image matching without relying on 3D supervision.
  • Implemented a labeling pipeline to avoid manual annotation and ensured coarse alignment via GPS.
  • Adopted bird's-eye-view transformations to address perspective distortion.
  • Ensembled multiple open-source matchers to generate and refine correspondences.
  • Achieved a correspondence coverage increase from 30.5% to 82.2%.
  • Demonstrated 98.6% accuracy in matching.
  • Maintained robustness across different camera types including smartphones and digital cameras.

Abstract

Vehicle-mounted smartphones are used to increase the frequency of road inspections. If images taken at different times can be accurately matched, temporal changes in road conditions become more apparent for damage analysis. However, the low quality of smartphone sensors, road deterioration, and lighting variations make matching unreliable. Existing matchers generalize poorly to road scenes and are usually trained with 3D supervision, such as depth maps and camera poses, which are unavailable for smartphones. To address this limitation, a self-supervised adaptation with pixel-level precision is proposed, in which the supervision of an existing dense matching model is redesigned to enable training solely on RGB images. A self-supervised labeling pipeline eliminates manual annotation of point correspondences. GPS ensures coarse alignment and bird’s-eye-view transformation mitigates perspective distortion. Then multiple open-source matchers are ensembled to propose correspondences, which are refined by robust fitting and propagated to unlabeled regions. Over five years, 76,882 image pairs were collected across 20 regions in Japan and China. The proposed method boosts correspondence coverage from 30.5% to 82.2% with 98.6% accuracy, maintaining robustness across smartphones, conventional digital cameras, and line-scan cameras.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Xue et al. (2026) studied this question.

synapsesocial.com/papers/69b4fbeab39f7826a300c5f8https://doi.org/10.1016/j.cacaie.2026.100027
Ask AI
Helpful
Bookmark
Share
View Full Paper