PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
August 13, 2026Scientific Reports0 citationsOpen Access

Benchmark of small language models for plain-language simplification in Spanish clinical texts

PMPaloma Martı́nezJSJesús M. Sánchez-GómezLMLourdes Moreno

Key Points

  • This work aims to establish a benchmark for small language models to simplify clinical texts in Spanish, enhancing health literacy.
  • Introduced MEDICLARO corpus with 50 clinical notes and three simplifications each from experts.
  • Evaluated four families of language models using fine-tuning and prompt-based strategies.
  • Assessed simplification performance through metrics like semantic similarity, factual consistency, and readability.
  • Llama-3.2-3B showed the best balance across efficiency, robustness, and overall performance.
  • RigoChat-v2-7B excelled in output quality and readability.
  • Findings establish a foundation for integrating small language models into Spanish clinical workflows.

Abstract

Abstract Medical text simplification is important for improving health literacy and making clinical information easier for patients to understand. While clinical text simplification has been widely studied in English, Spanish remains underexplored, especially for systematic adaptation into plain-language in clinical settings. This work presents the first benchmark of small language models for Spanish clinical plain-language adaptation and introduces MEDICLARO, a corpus specifically designed for this task. MEDICLARO consists of 50 clinical notes, each with three human-written simplifications produced by cognitive accessibility experts in accordance with ISO 24495-1:2023. Four families of state-of-the-art language models were evaluated through fine-tuning and prompt-based strategies. The evaluation covers simplification, semantic similarity, factual consistency, readability, and environmental impact, and is complemented by human evaluation and qualitative error analysis. The results show that Llama-3.2-3B provides the most balance profile across efficiency, robustness, and overall performance, while RigoChat-v2-7B stands out when output quality and readability are prioritized. Overall, this work establishes a solid foundation for integrating small language models into Spanish clinical workflows, offering a sustainable, patient-centered path toward digital health accessibility.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Martı́nez et al. (2026) studied this question.

synapsesocial.com/papers/6a7d76792b0e0cff3f63fc88https://doi.org/10.1038/s41598-026-65740-w
Ask AI
Helpful
Bookmark
Share
View Full Paper