PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 25, 2026The Journal of Supercomputing0 citationsOpen Access

CNN-TriFuseResNet-50: a hybrid deep learning model for efficient and interpretable detection of middle cerebral artery occlusion

KZKun ZhuNSNur Hana SamsudinHQHang Qu

Key Points

  • To develop a hybrid deep learning model for detecting middle cerebral artery occlusion in acute ischemic stroke using non-contrast CT images.
  • Developed CNN-TriFuseResNet-50 architecture incorporating triplet attention and lightweight convolutional blocks.
  • Evaluated performance on a retrospective cohort of 822 cases of middle cerebral artery occlusion without hyperdense artery sign.
  • Compared accuracy and parameter efficiency with Vision Transformers on the same dataset.
  • Achieved an accuracy of 81.66% and AUC of 0.90, outperforming Vision Transformers by 4.37%.
  • Reduced false negatives by 24%, achieving 84.52% sensitivity.
  • Demonstrated real-time inference capabilities at 23 frames per second.

Abstract

Acute ischemic stroke (AIS) is a time-critical medical emergency where high-performance computing (HPC) capabilities are essential to enable real-time diagnosis and rapid treatment decision-making. Acute ischemic stroke caused by middle cerebral artery occlusion (MCAO) without hyperdense artery sign (HAS) on non-contrast CT (NCCT) poses significant diagnostic challenges due to subtle imaging features. To address this, we propose CNN-TriFuseResNet-50, a novel deep learning architecture integrating Triplet Attention, pre-trained ResNet-50, and lightweight convolutional blocks for automated HAS-negative MCAO detection. The Triplet Attention mechanism concurrently models spatial (width, height) and channel-wise dependencies, resolving feature conflicts inherent in conventional attention modules like Squeeze-and-Excitation (SE) and Convolutional Block Attention Module (CBAM). Evaluated on a retrospective cohort of 822 cases, our architecture achieves 81.66% accuracy and 0.90 AUC, outperforming Vision Transformers (ViT) by 4.37% accuracy with 58% fewer parameters. This architecture reduces false negatives by 24% (84.52% sensitivity) and achieves real-time inference (23 frames per second FPS), satisfying the stringent latency requirements of high-performance medical imaging systems. Class activation maps (CAMs) provide interpretability by localizing occlusion regions aligned with angiographic ground truth. This work bridges the gap between computational efficiency and diagnostic accuracy, offering a deployable solution for early stroke management in resource-constrained settings.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Zhu et al. (2026) studied this question.

synapsesocial.com/papers/699e90eff5123be5ed04e263https://doi.org/10.1007/s11227-026-08311-0
Ask AI
Helpful
Bookmark
Share
View Full Paper