Tracing Back Music Emotion Predictions to Sound Sources and Intuitive   Perceptual Qualities

Shreyan Chowdhury; Verena Praher; Gerhard Widmer

arXiv:2106.07787·cs.SD·June 17, 2021·6 cites

Tracing Back Music Emotion Predictions to Sound Sources and Intuitive Perceptual Qualities

Shreyan Chowdhury, Verena Praher, Gerhard Widmer

PDF

Open Access 2 Repos

TL;DR

This paper introduces a method combining source separation and perceptual features to interpret music emotion predictions, enhancing understanding of model decisions and aiding in debugging biases.

Contribution

It presents an innovative approach that merges audioLIME with perceptual features, providing more intuitive explanations for music emotion recognition models.

Findings

01

Improved interpretability of emotion prediction models.

02

Effective debugging of biased models.

03

Enhanced connection between audio inputs and emotion outputs.

Abstract

Music emotion recognition is an important task in MIR (Music Information Retrieval) research. Owing to factors like the subjective nature of the task and the variation of emotional cues between musical genres, there are still significant challenges in developing reliable and generalizable models. One important step towards better models would be to understand what a model is actually learning from the data and how the prediction for a particular input is made. In previous work, we have shown how to derive explanations of model predictions in terms of spectrogram image segments that connect to the high-level emotion prediction via a layer of easily interpretable perceptual features. However, that scheme lacks intuitive musical comprehensibility at the spectrogram level. In the present work, we bridge this gap by merging audioLIME -- a source-separation based explainer -- with mid-level…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsMusic and Audio Processing · Neuroscience and Music Perception · Neural Networks and Applications