PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 26, 2026Electronics1 citationsOpen Access

YOLO11-MSCA: A Multi-Scale Channel Attention Model for Lumbar Vertebra Detection in X-Ray Images

View Full Paper
HFHana Ben FredjHGHatem GarrabCSChokri Souani

Key Points

  • The research aims to enhance the automated detection of lumbar vertebrae in X-ray images using a novel attention model.
  • Integration of Multi-Scale Channel Attention Block within the YOLO11 backbone
  • Training on AP-view images from the Burapha University Lumbar-Spine Dataset
  • End-to-end model operation without complex preprocessing
  • Achieved a mean Average Precision (mAP) of 0.982 at IoU 0.5 and 0.79 at IoU 0.5–0.95
  • Precision of 0.93 and recall of 0.975
  • Demonstrated improved detection performance over the YOLO11 baseline with minimal increase in model complexity

Abstract

Automated identification of lumbar vertebrae plays a key role in modern spine analysis, offering valuable assistance for diagnostic assessment and preoperative decision-making. Despite recent progress in deep learning-based detection methods, accurately localizing vertebral structures remains challenging due to anatomical variability and heterogeneous image quality. To address the difficulty of capturing subtle vertebral structures, we introduce a Multi-Scale Channel Attention Block (MSCABlock) integrated into the YOLO11 backbone. Unlike conventional attention-based or multi-scale convolutional designs, MSCABlock jointly exploits channel-wise feature interaction and multi-scale receptive fields to enhance both local detail sensitivity and contextual representation, while preserving computational efficiency. The proposed approach is designed to improve detection performance without significantly increasing model complexity. Our model is trained and validated using only the AP-view images from the Burapha University Lumbar-Spine Dataset (BUU-LSPINE), which provides well-annotated lumbar spine X-ray images from 400 unique patients. The proposed approach operates in a fully end-to-end manner, allowing vertebrae to be identified directly from input images without relying on handcrafted feature engineering or complex preprocessing pipelines. Experimental evaluations show that the proposed model achieves strong detection performance, with mAP@0.5 and mAP@0.5–0.95 reaching 0.982 and 0.79, respectively, alongside a precision of 0.93 and a recall of 0.975. Compared with the YOLO11 baseline, ablation and efficiency analyses demonstrate that MSCABlock consistently improves detection performance. It introduces only marginal increases in model parameters and computational cost, thereby preserving a lightweight architecture and maintaining efficient inference. These results show that the optimized YOLO11-based system generalizes well across lumbar levels. It maintains reliable detection under challenging conditions, providing robust automated localization to support large-scale clinical spine analysis.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Fredj et al. (2026) studied this question.

synapsesocial.com/papers/69c4ccaffdc3bde4489180e6https://doi.org/10.3390/electronics15071341
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Going deeper with convolutions2015 · 47,236 citations
  2. 2Exploring Neighbor Spatial Relationships for Enhanced Lumbar Vertebrae Detection in X-ray Images2024 · 3 citations
  3. 3Deep Learning-Based Automated Magnetic Resonance Image Segmentation of the Lumbar Structure and Its Adjacent Structures at the L4/5 Level2023 · 6 citations
  4. 4Automated detection and classification of acute vertebral body fractures using a convolutional neural network on computed tomography2023 · 47 citations
  5. 5Identification of L5 vertebra on lumbar spine radiographs using deep learning2024 · 7 citations