PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 29, 2026PLoS ONE1 citationsOpen Access

MedZeroSeg: Zero-shot medical image segmentation via vision foundation models

View Full Paper
RZRonghui ZhangMHMin HuangRLRui Li

Key Points

  • The study aims to develop a framework for zero-shot medical image segmentation using advanced vision models.
  • Introduced MedZeroSeg for medical image segmentation challenges.
  • Utilized CLIP and SAM for enhanced zero-shot capabilities.
  • Implemented a Dual-Path Feature Extraction Module for detailed and contextual understanding.
  • Developed a Context-Enhanced Hard-Negative Contrast Loss for improved contrastive learning.
  • Conducted experiments on cardiac MRI, multi-organ abdominal CT, and chest X-ray datasets.
  • MedZeroSeg demonstrated superior performance in zero-shot segmentation tasks.
  • Achieved strong generalization across different medical imaging modalities.
  • Showed minimal dependency on large annotated datasets.

Abstract

A novel medical image segmentation framework, MedZeroSeg , is proposed to address key challenges in the field. Leveraging vision foundation models such as CLIP (Contrastive Language-Image Pre-training) and SAM (Segment Anything Model), it achieves zero-shot segmentation, accurately delineating previously unseen medical images without requiring additional labeled data. This significantly reduces reliance on large-scale annotated datasets. At its core, MedZeroSeg introduces a Dual-Path Feature Extraction Module that captures both fine anatomical details and global contextual information through the integration of local and global perception mechanisms, enhancing robustness against the complexity and variability inherent in medical imaging.Additionally, a Context-Enhanced Hard-Negative Contrast Loss is introduced to enhance contrastive learning by exploiting contextual cues and refining hard-negative sampling, leading to better representations and higher efficiency. The key innovation of MedZeroSeg lies in its ability to leverage generalizable knowledge from CLIP and SAM without any task-specific fine-tuning, making it highly adaptable across different medical imaging modalities. Extensive experiments on three publicly available datasets, including cardiac MRI (ACDC), multi-organ abdominal CT (Synapse), and chest X-ray (COVID-QU-Ex), demonstrate that MedZeroSeg achieves superior results in both zero-shot and weakly supervised segmentation settings, showcasing strong generalization capabilities and minimal data dependency. The framework represents a significant advancement in medical image analysis and opens up promising directions for future research in applying advanced foundation models and innovative learning strategies to healthcare applications.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Zhang et al. (2026) studied this question.

synapsesocial.com/papers/69c8c3cede0f0f753b39ecf5https://doi.org/10.1371/journal.pone.0344978
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Grad-CAM: Visual Explanations from Deep Networks via Gradient-Based Localization2017 · 22,844 citations
  2. 2Mutual Information Guided Diffusion for Zero-Shot Cross-Modality Medical Image Translation2024 · 50 citations
  3. 3Radiology Objects in COntext (ROCO): A Multimodal Image Dataset2018 · 193 citations
  4. 4U-Net: Convolutional Networks for Biomedical Image Segmentation2015 · 92,136 citations
  5. 5SurgicalGaussian: Deformable 3D Gaussians for High-Fidelity Surgical Scene Reconstruction2024 · 16 citations