An Ensemble Approach to Music Source Separation: A Comparative Analysis of Conventional and Hierarchical Stem Separation
Saarth Vardhan, Pavani R Acharya, Samarth S Rao, Oorjitha Ratna Jasthi, and S Natarajan

TL;DR
This paper introduces an ensemble method for music source separation that combines multiple models to improve the isolation of vocal, drum, and bass stems, and explores hierarchical separation into sub-stems, revealing insights into factors affecting performance.
Contribution
The paper presents a novel ensemble approach that leverages multiple architectures for improved MSS and extends into hierarchical sub-stem separation, addressing limitations of single-model methods.
Findings
Ensemble approach outperforms individual models on VDB stems.
Hierarchical separation successfully isolates sub-stems like kick and snare.
Performance varies with genre and instrumentation, highlighting complexity.
Abstract
Music source separation (MSS) is a task that involves isolating individual sound sources, or stems, from mixed audio signals. This paper presents an ensemble approach to MSS, combining several state-of-the-art architectures to achieve superior separation performance across traditional Vocal, Drum, and Bass (VDB) stems, as well as expanding into second-level hierarchical separation for sub-stems like kick, snare, lead vocals, and background vocals. Our method addresses the limitations of relying on a single model by utilising the complementary strengths of various models, leading to more balanced results across stems. For stem selection, we used the harmonic mean of Signal-to-Noise Ratio (SNR) and Signal-to-Distortion Ratio (SDR), ensuring that extreme values do not skew the results and that both metrics are weighted effectively. In addition to consistently high performance across the…
Peer Reviews
No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.
Videos
No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.
Taxonomy
TopicsMusic and Audio Processing · Music Technology and Sound Studies · Diverse Musicological Studies
