PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 28, 20260 citationsOpen Access

Performance Comparison of AI Models for Image Shadow Removal: UNet, CGAN, and Swin-Transformer with a Note on Diffusion Models

SZShangan Zhou

Key Points

  • This analysis aims to compare the effectiveness of UNet, CGAN, and Swin-Transformer in removing shadows from images.
  • Comparison of UNet, CGAN, and Swin-Transformer models on the ISTD benchmark dataset.
  • Evaluation using quantitative metrics: PSNR, SSIM, RMSE, MAE.
  • Qualitative visual assessment of shadow removal performance.
  • Swin-Transformer outperforms other models in detail preservation and artifact reduction.
  • CGAN demonstrates enhanced perceptual realism.
  • UNet offers a computationally efficient baseline for image shadow removal applications.

Abstract

This study conducts a comprehensive performance comparison of three prominent deep learning architectures—UNet, Conditional Generative Adversarial Network (CGAN), and Swin-Transformer—for the task of single-image shadow removal, with additional theoretical consideration given to Denoising Diffusion Probabilistic Models (DDPM). Evaluated on the ISTD benchmark dataset using quantitative metrics (PSNR, SSIM, RMSE, MAE) and qualitative visual assessment, the results establish a clear performance hierarchy. The Swin-Transformer model consistently achieves superior results, excelling in detail preservation, artifact reduction, and maintaining global illumination consistency, attributed to its hierarchical structure and shifted-window self-attention mechanism. The CGAN model demonstrates enhanced perceptual realism through adversarial training, while the UNet provides a computationally efficient baseline. The findings offer practical guidance for model selection based on specific application requirements and highlight the impact of architectural design. This analysis concludes by suggesting future research pathways, including the exploration of hybrid models and the empirical application of diffusion models for high-fidelity image restoration tasks.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Shangan Zhou (2026) studied this question.

synapsesocial.com/papers/69a287690a974eb0d3c03263https://doi.org/10.5281/zenodo.18786374
Ask AI
Helpful
Bookmark
Share
View Full Paper