PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 18, 20242 citationsOpen Access

Unsupervised Pitch-Timbre Disentanglement of Musical Instruments Using a Jacobian Disentangled Sequential Autoencoder

View Full Paper
YLYin-Jyun LuoSESebastian EwertSDSimon Dixon

Key Points

Key points are not available for this paper at this time.

Abstract

Disentangled representation learning seeks to align individual dimensions or separate groups of coordinates of latent factors with attributes of observed data such that perturbing certain latent factors uniquely changes particular attributes. A main challenge in unsupervised disentanglement using autoencoders is that strong regularisation, while necessary for consistent disentanglement, comes at the expense of accurate data reconstruction. To address this, we introduce a teacher-student framework that incorporates a variational sequential autoencoder and a Jacobian constraint that regularises the variation of observations relative to latent factors. In real-world audio recordings of musical instruments, our approach outperforms a state-of-the-art method in both sampling quality and unsupervised pitch-timbre disentanglement.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Luo et al. (2024) studied this question.

synapsesocial.com/papers/68e7398bb6db6435876b2c29https://doi.org/10.1109/icassp48485.2024.10447564
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 3 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Fréchet Audio Distance: A Reference-Free Metric for Evaluating Music Enhancement Algorithms2019 · 187 citations
  2. 2Melody Extraction from Polyphonic Music Signals: Approaches, applications, and challenges2014 · 232 citations
  3. 3Creating a Multitrack Classical Music Performance Dataset for Multimodal Music Analysis: Challenges, Insights, and Applications2018 · 136 citations