PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
June 16, 20243 citationsOpen Access

SingMOS: An extensive Open-Source Singing Voice Dataset for MOS Prediction

View Full Paper
YTYuxun TangJSJiatong ShiYWYuning Wu

Key Points

Key points are not available for this paper at this time.

Abstract

In speech generation tasks, human subjective ratings, usually referred to as the opinion score, are considered the "gold standard" for speech quality evaluation, with the mean opinion score (MOS) serving as the primary evaluation metric. Due to the high cost of human annotation, several MOS prediction systems have emerged in the speech domain, demonstrating good performance. These MOS prediction models are trained using annotations from previous speech-related challenges. However, compared to the speech domain, the singing domain faces data scarcity and stricter copyright protections, leading to a lack of high-quality MOS-annotated datasets for singing. To address this, we propose SingMOS, a high-quality and diverse MOS dataset for singing, covering a range of Chinese and Japanese datasets. These synthesized vocals are generated using state-of-the-art models in singing synthesis, conversion, or resynthesis tasks and are rated by professional annotators alongside real vocals. Data analysis demonstrates the diversity and reliability of our dataset. Additionally, we conduct further exploration on SingMOS, providing insights for singing MOS prediction and guidance for the continued expansion of SingMOS.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Tang et al. (2024) studied this question.

synapsesocial.com/papers/68e64877b6db6435875d974ahttps://doi.org/10.48550/arxiv.2406.10911
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1SingMOS-Pro: An Comprehensive Benchmark For Singing Quality Assessment2026
  2. 2CodecMOS: Singing MOS prediction through the integration of self-supervised speech representations and neural audio codec features2026
  3. 3From Scores to Preferences: Redefining MOS Benchmarking for Speech Quality Reward Modeling2025
  4. 4SingNet: Towards a Large-Scale, Diverse, and In-the-Wild Singing Voice Dataset2025
  5. 5Singing Voice Data Scaling-up: An Introduction to ACE-Opencpop and ACE-KiSing2024 · 9 citations