PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 21, 2026Epilepsy Research2 citationsOpen Access

Independent Evaluation of Deep Learning Models for Detecting Focal Cortical Dysplasia

View Full Paper
HKHelene KaasMPMartin PrenerMGMelanie Ganz

Key Points

  • This study aims to evaluate the effectiveness of three deep learning models in detecting focal cortical dysplasia using MRI images.
  • Assessed three deep learning tools: DeepFCD, MELD Classifier, and MELDGraph.
  • Included 101 epilepsy patients with FCD and 101 without FCD from MRI scans.
  • Evaluated classifiers at both patient-level and lesion-level for FCD detection.
  • Calculated test-retest consistency using the Dice coefficient from repeated MRI scans.
  • MELDClassifier achieved 52% accuracy; MELDGraph reached 61%; DeepFCD performed at 56%.
  • At lesion-level, MELDClassifier had a sensitivity of 70% and PPV of 13%; MELDGraph had 53% sensitivity and PPV of 36%; DeepFCD showed 30% sensitivity and PPV of 19%.
  • Average Dice coefficients for test-retest reliability ranged from 0.28 to 0.38 across classifiers.

Abstract

The purpose of this study is to perform an independent assessment of three state-of-the-art tools for the detection of focal cortical dysplasia (FCD) from Magnetic Resonance images (MRI). These tools include DeepFCD, the Multi-center Epilepsy Lesion Detection (MELD) Classifier, and MELDGraph. T1-weighted and fluid-attenuated inversion recovery MRIs from 101 epilepsy patients with FCD and 101 epilepsy patients without FCD were retrospectively included. Classifiers were evaluated at patient-level by their ability to correctly identify the presence of any FCD lesions, and at lesion-level by their capacity to identify lesions within regions delineated by neuroradiologists in MRI reports. A calibrated threshold for DeepFCD prediction probabilities was empirically determined to improve classifier specificity. Classifier test-retest consistency was measured using the Dice coefficient on repeated MRI scans of 21 individuals. At patient-level, MELDClassifier achieved 52% accuracy (sensitivity=91%, specificity=14%), MELDGraph reached 61% accuracy (sensitivity=76%, specificity=47%) and DeepFCD performed with 56% accuracy (sensitivity=62%, specificity=50%) at an empirically determined threshold of 0.90. At lesion-level, MELDClassifier performed with a sensitivity of 70% and a positive predictive value (PPV) of 13%. MELDGraph reached 53% sensitivity and PPV of 36%, whereas the DeepFCD performed with 30% sensitivity and PPV of 19%. Test-retest reliability was low, with an average min, max Dice coefficient of 0.28 0.0, 1.0 for MELDClassifier, 0.38 0.0, 1.0 for MELDGraph, and 0.35 0.05, 0.54 for DeepFCD. This study highlights the current limitations of using deep learning models in FCD diagnosis and emphasizes the need to enhance the tools’ accuracy, reliability, and interpretability to improve clinical utility. • We present the largest test dataset of consistent scanner and sequence parameters • The accuracy of deep learning tools detecting focal cortical dysplasia is 52-61% • Test-retest of repeated MRI scans revealed average Dice coefficients of 0.28-0.38 • At lesion-level, average false positive count per patient ranged from 0.49-32.71 • Better accuracy, reliability, and interpretability will improve clinical utility

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Kaas et al. (2026) studied this question.

synapsesocial.com/papers/69be38906e48c4981c679120https://doi.org/10.1016/j.eplepsyres.2026.107778
Ask AI
Helpful
Bookmark
Share
View Full Paper