PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
October 16, 20250 citationsOpen Access

Fairness in Dysarthric Speech Synthesis: Understanding Intrinsic Bias in Dysarthric Speech Cloning using F5-TTS

View Full Paper
MAM AnuprabhaKGKrishna GurugubelliAVAnil Kumar Vuppala

Key Points

  • F5-TTS exhibits a strong bias towards intelligibility in dysarthric speech synthesis, impacting speaker similarity.
  • Using the TORGO dataset, we evaluated speaker similarity, prosody preservation, and intelligibility metrics relevant to dysarthric speech.
  • Fairness metrics like Disparate Impact were applied to assess speech synthesis across different dysarthric severity levels.
  • Insights from this analysis can drive more inclusive speech technologies by addressing biases in dysarthric speech synthesis.

Abstract

Dysarthric speech poses significant challenges in developing assistive technologies, primarily due to the limited availability of data. Recent advances in neural speech synthesis, especially zero-shot voice cloning, facilitate synthetic speech generation for data augmentation; however, they may introduce biases towards dysarthric speech. In this paper, we investigate the effectiveness of state-of-the-art F5-TTS in cloning dysarthric speech using TORGO dataset, focusing on intelligibility, speaker similarity, and prosody preservation. We also analyze potential biases using fairness metrics like Disparate Impact and Parity Difference to assess disparities across dysarthric severity levels. Results show that F5-TTS exhibits a strong bias toward speech intelligibility over speaker and prosody preservation in dysarthric speech synthesis. Insights from this study can help integrate fairness-aware dysarthric speech synthesis, fostering the advancement of more inclusive speech technologies.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Anuprabha et al. (2025) studied this question.

synapsesocial.com/papers/68f12bfb2107091eab27a3efhttps://doi.org/10.48550/arxiv.2508.05102
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Voice Cloning for Dysarthric Speech Synthesis: Addressing Data Scarcity in Speech-Language Pathology2025
  2. 2Improved Dysarthric Speech to Text Conversion via TTS Personalization2025
  3. 3What determines the success of AI voice-cloned speech? Prosodic and acoustic evidence on three TTS systems2026
  4. 4Exploring Dysphonic Artificial Intelligence Voice Cloning for Speech Intelligibility in Noise2026
  5. 5Personalized Text to Speech Synthesis through Few Shot Speaker Adaptation with Contrastive Learning2025