TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Papers/Music Source Separation with Band-split RNN

Music Source Separation with Band-split RNN

Yi Luo, Jianwei Yu

2022-09-30Music Source Separation
PaperPDFCodeCodeCode

Abstract

The performance of music source separation (MSS) models has been greatly improved in recent years thanks to the development of novel neural network architectures and training pipelines. However, recent model designs for MSS were mainly motivated by other audio processing tasks or other research fields, while the intrinsic characteristics and patterns of the music signals were not fully discovered. In this paper, we propose band-split RNN (BSRNN), a frequency-domain model that explictly splits the spectrogram of the mixture into subbands and perform interleaved band-level and sequence-level modeling. The choices of the bandwidths of the subbands can be determined by a priori knowledge or expert knowledge on the characteristics of the target source in order to optimize the performance on a certain type of target musical instrument. To better make use of unlabeled data, we also describe a semi-supervised model finetuning pipeline that can further improve the performance of the model. Experiment results show that BSRNN trained only on MUSDB18-HQ dataset significantly outperforms several top-ranking models in Music Demixing (MDX) Challenge 2021, and the semi-supervised finetuning stage further improves the performance on all four instrument tracks.

Results

TaskDatasetMetricValueModel
Music Source SeparationMUSDB18SDR (avg)8.97Band-Split RNN (semi-sup.)
Music Source SeparationMUSDB18SDR (bass)8.16Band-Split RNN (semi-sup.)
Music Source SeparationMUSDB18SDR (drums)10.15Band-Split RNN (semi-sup.)
Music Source SeparationMUSDB18SDR (other)7.08Band-Split RNN (semi-sup.)
Music Source SeparationMUSDB18SDR (vocals)10.47Band-Split RNN (semi-sup.)
Music Source SeparationMUSDB18SDR (avg)8.23Band-Split RNN
Music Source SeparationMUSDB18SDR (bass)7.51Band-Split RNN
Music Source SeparationMUSDB18SDR (drums)8.58Band-Split RNN
Music Source SeparationMUSDB18SDR (other)6.62Band-Split RNN
Music Source SeparationMUSDB18SDR (vocals)10.21Band-Split RNN
Music Source SeparationMUSDB18-HQSDR (avg)8.97Band-Split RNN (semi-sup.)
Music Source SeparationMUSDB18-HQSDR (bass)8.16Band-Split RNN (semi-sup.)
Music Source SeparationMUSDB18-HQSDR (drums)10.15Band-Split RNN (semi-sup.)
Music Source SeparationMUSDB18-HQSDR (others)7.08Band-Split RNN (semi-sup.)
Music Source SeparationMUSDB18-HQSDR (vocals)10.47Band-Split RNN (semi-sup.)
Music Source SeparationMUSDB18-HQSDR (avg)8.24Band-Split RNN
Music Source SeparationMUSDB18-HQSDR (bass)7.22Band-Split RNN
Music Source SeparationMUSDB18-HQSDR (drums)9.01Band-Split RNN
Music Source SeparationMUSDB18-HQSDR (others)6.7Band-Split RNN
Music Source SeparationMUSDB18-HQSDR (vocals)10.01Band-Split RNN
2D ClassificationMUSDB18SDR (avg)8.97Band-Split RNN (semi-sup.)
2D ClassificationMUSDB18SDR (bass)8.16Band-Split RNN (semi-sup.)
2D ClassificationMUSDB18SDR (drums)10.15Band-Split RNN (semi-sup.)
2D ClassificationMUSDB18SDR (other)7.08Band-Split RNN (semi-sup.)
2D ClassificationMUSDB18SDR (vocals)10.47Band-Split RNN (semi-sup.)
2D ClassificationMUSDB18SDR (avg)8.23Band-Split RNN
2D ClassificationMUSDB18SDR (bass)7.51Band-Split RNN
2D ClassificationMUSDB18SDR (drums)8.58Band-Split RNN
2D ClassificationMUSDB18SDR (other)6.62Band-Split RNN
2D ClassificationMUSDB18SDR (vocals)10.21Band-Split RNN
2D ClassificationMUSDB18-HQSDR (avg)8.97Band-Split RNN (semi-sup.)
2D ClassificationMUSDB18-HQSDR (bass)8.16Band-Split RNN (semi-sup.)
2D ClassificationMUSDB18-HQSDR (drums)10.15Band-Split RNN (semi-sup.)
2D ClassificationMUSDB18-HQSDR (others)7.08Band-Split RNN (semi-sup.)
2D ClassificationMUSDB18-HQSDR (vocals)10.47Band-Split RNN (semi-sup.)
2D ClassificationMUSDB18-HQSDR (avg)8.24Band-Split RNN
2D ClassificationMUSDB18-HQSDR (bass)7.22Band-Split RNN
2D ClassificationMUSDB18-HQSDR (drums)9.01Band-Split RNN
2D ClassificationMUSDB18-HQSDR (others)6.7Band-Split RNN
2D ClassificationMUSDB18-HQSDR (vocals)10.01Band-Split RNN

Related Papers

Music Source Restoration2025-05-27Training-Free Multi-Step Audio Source Separation2025-05-26Is MixIT Really Unsuitable for Correlated Sources? Exploring MixIT for Unsupervised Pre-training in Music Source Separation2025-05-12Solving Copyright Infringement on Short Video Platforms: Novel Datasets and an Audio Restoration Deep Learning Pipeline2025-04-30Score-informed Music Source Separation: Improving Synthetic-to-real Generalization in Classical Music2025-03-10Separate This, and All of these Things Around It: Music Source Separation via Hyperellipsoidal Queries2025-01-27Sanidha: A Studio Quality Multi-Modal Dataset for Carnatic Music2025-01-12MAJL: A Model-Agnostic Joint Learning Framework for Music Source Separation and Pitch Estimation2025-01-07