PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
September 28, 2025Journal of Artificial Intelligence and Soft Computing Research3 citationsOpen Access

Chestxgen: Dynamic Memory-Augmented Vision-Language Transformer with Context-Aware Gating for Radiology Report Generation

View Full Paper
SASharofiddin AllaberdievAKAsif KhanSMSardor Mamarasulov

Key Points

  • ChestXGen achieves significant improvements in report generation accuracy for radiology, reducing the burden on radiologists.
  • BLEU and METEOR metric evaluations on the MIMIC-CXR dataset showcase performance enhancements compared to previous models.
  • The integrated framework leverages Transformer architecture along with memory augmentation to better handle rare disease detection.
  • Significant results suggest that automated report generation tools like ChestXGen may enhance the overall quality of diagnostic evaluations.

Abstract

Abstract Chest X-ray analysis is vital for clinical screening, diagnosis, and treatment planning. The increasing workload on radiologists calls for robust automated solutions to generate accurate and standardized reports. Conventional report generation models often struggle to detect rare and anomalous diseases, particularly when faced with imbalanced datasets, which can compromise diagnostic knowledge accuracy. To address these limitations, we propose ChestXGen, a novel multimodal framework for automated radiology report generation. Our model is based on a fully Transformer-based encoder-decoder architecture that integrates Memory Augmented Transformer (MAT) blocks with a Context-Aware Bi-Gate (CABG) mechanism. These enable the model to capture long-range dependencies, effectively fuse visual and textual features, and better handle underrepresented conditions. Visual features are extracted using a ResNet-101-V2 backbone and refined through a shared memory module that continuously reinforces cross-modal associations. This integrated approach facilitates the generation of comprehensive, accurate, and contextually coherent reports. Extensive evaluation on the large-scale MIMIC-CXR dataset, comprising 377,110 images and corresponding free-text reports demonstrate that ChestXGen outperforms previous models on BLEU-1, BLEU-2, BLEU-3, and METEOR metrics. The results demonstrate the efficacy of Transformer-based models in substantially reducing radiologists’ reporting burden while concurrently enhancing the precision and reliability of diagnostic interpretations.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Allaberdiev et al. (2025) studied this question.

synapsesocial.com/papers/68d8f313d88e2624dc4c56a5https://doi.org/10.2478/jaiscr-2026-0003
Ask AI
Helpful
Bookmark
Share
View Full Paper