PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
October 8, 2025Remote Sensing3 citationsOpen Access

LMVMamba: A Hybrid U-Shape Mamba for Remote Sensing Segmentation with Adaptation Fine-Tuning

View Full Paper
LFLi FanXWXiao WangHWHaochen Wang

Key Points

  • LMVMamba achieves superior results in semantic segmentation, enhancing both global and local feature recognition.
  • On the OpenEarthMap dataset, LMVMamba reached mIoU of 52.3% and OA of 69.8%, showcasing its performance.
  • Analysis employed a neural network with LoRA fine-tuning and a U-shaped Mamba decoder to improve segmentation precision.
  • The findings underline the model's capacity for high-precision land-cover classification in remote sensing applications.

Abstract

High-precision semantic segmentation of remote sensing imagery is crucial in geospatial analysis. It plays an immeasurable role in fields such as urban governance, environmental monitoring, and natural resource management. However, when confronted with complex objects (such as winding roads and dispersed buildings), existing semantic segmentation methods still suffer from inadequate target recognition capabilities and multi-scale representation issues. This paper proposes a neural network model, LMVMamba (LoRA Multi-scale Vision Mamba), for semantic segmentation of remote sensing images. This model integrates the advantages of convolutional neural networks (CNNs), Transformers, and state-space models (Mamba) with a multi-scale feature fusion strategy. It simultaneously captures global contextual information and fine-grained local features. Specifically, in the encoder stage, the ResT Transformer serves as the backbone network, employing a LoRA fine-tuning strategy to effectively enhance model accuracy by training only the introduced low-rank matrix pairs. The extracted features are then passed to the decoder, where a U-shaped Mamba decoder is designed. In this stage, a Multi-Scale Post-processing Block (MPB) is introduced, consisting of depthwise separable convolutions and residual concatenation. This block effectively extracts multi-scale features and enhances local detail extraction after the VSS block. Additionally, a Local Enhancement and Fusion Attention Module (LAS) is added at the end of each decoder block. LAS integrates the SimAM attention mechanism, further enhancing the model’s multi-scale feature fusion capability and local detail segmentation capability. Through extensive comparative experiments, it was found that LMVMamba achieves superior performance on the OpenEarthMap dataset (mIoU 52.3%, OA 69.8%, mF1: 68.0%) and LoveDA (mIoU 67.9%, OA 80.3%, mF1: 80.5%) datasets. Ablation experiments validated the effectiveness of each module. The final results indicate that this model is highly suitable for high-precision land-cover classification tasks in remote sensing imagery. LMVMamba provides an effective solution for precise semantic segmentation of high-resolution remote sensing imagery.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Fan et al. (2025) studied this question.

synapsesocial.com/papers/68e5c1b46950a706b22b50ebhttps://doi.org/10.3390/rs17193367
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Semantic segmentation of high spatial resolution images with deep neural networks2019 · 38 citations
  2. 2MobileMamba: Lightweight Multi-Receptive Visual Mamba Network2025 · 84 citations
  3. 3Effective Sequential Classifier Training for SVM-Based Multitemporal Remote Sensing Image Classification2018 · 121 citations
  4. 4A survey of remote sensing image classification based on CNNs2019 · 234 citations
  5. 5A novel image fusion method based on UAV and Sentinel-2 for environmental monitoring2025 · 18 citations