PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
April 18, 2026BMC Oral Health0 citationsOpen Access

Comparative analysis of transformer, CNN, and YOLO architectures for mandibular condyle segmentation on panoramic radiographs: a deep learning benchmark

SYSerkan YılmazSÖSerdar ÖztürkHÖHatice Seda Özgedik

Key Points

  • The aim is to compare the performance of different deep learning architectures for mandibular condyle segmentation.
  • Dataset of 1,300 panoramic radiographs was curated for analysis.
  • Ground-truth masks annotated by a primary radiologist and reviewed for accuracy.
  • Six different architectures assessed for segmentation performance using metrics such as IoU and DSC.
  • Models evaluated on a fixed test set.
  • All models achieved high segmentation accuracy with DSC values between 0.819 and 0.866.
  • RT-DETR model had the highest DSC of 0.866 and IoU of 0.764.
  • YOLOv9-Seg provided competitive results with a DSC of 0.862 and high recall of 0.902.
  • YOLOv11-Seg showed high sensitivity but lower precision compared to others.

Abstract

This study aimed to perform the first multi-architecture comparison of pixel-level mandibular condyle segmentation on panoramic radiographs using transformer-based (RT-DETR), CNN-based (EfficientNet, Mask R-CNN, ConvNeXt), and YOLO-based (YOLOv9-Seg, YOLOv11-Seg) deep learning models. A dataset of 1,300 panoramic radiographs (2,600 condyles) was retrospectively curated. Ground-truth masks were annotated by a primary radiologist and reviewed by a senior radiologist; inter-observer agreement was quantified on a blinded 10% subset (Dice: 0.92 ± 0.03). Six state-of-the-art architectures were trained and evaluated on a fixed test set. Performance was assessed using Intersection over Union (IoU), Dice Similarity Coefficient (DSC), precision, recall, and F1-score. All models achieved high segmentation accuracy, with DSC values ranging from 0.819 to 0.866. The transformer-based RT-DETR model showed the highest numerical DSC (0.866), IoU (0.764), and F1-score (0.866), indicating a balanced overall segmentation profile. Among the one-stage detectors, YOLOv9-Seg provided competitive results (DSC: 0.862) with high recall (0.902), outperforming CNN-based alternatives. YOLOv11-Seg showed high sensitivity but lower precision compared to other architectures. Deep learning enables accurate and automated condylar segmentation on panoramic radiographs. While RT-DETR showed favorable anatomical fidelity for quantitative morphometry, YOLOv9-Seg presented a viable real-time alternative. This study establishes a benchmark for selecting segmentation architectures tailored to specific clinical needs in TMJ analysis. Not applicable.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Yılmaz et al. (2026) studied this question.

synapsesocial.com/papers/69e320af40886becb653fbc1https://doi.org/10.1186/s12903-026-08228-3
Ask AI
Helpful
Bookmark
Share
View Full Paper