{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/towards-unified-music-emotion-recognition","title":"Towards Unified Music Emotion Recognition across Dimensional and Categorical Models","arxiv_id":"2502.03979","date":"2025-02-06","proceeding":null,"authors":["Jaeyong Kang","Dorien Herremans"],"abstract":"One of the most significant challenges in Music Emotion Recognition (MER) comes from the fact that emotion labels can be heterogeneous across datasets with regard to the emotion representation, including categorical (e.g., happy, sad) versus dimensional labels (e.g., valence-arousal). In this paper, we present a unified multitask learning framework that combines these two types of labels and is thus able to be trained on multiple datasets. This framework uses an effective input representation that combines musical features (i.e., key and chords) and MERT embeddings. Moreover, knowledge distillation is employed to transfer the knowledge of teacher models trained on individual datasets to a student model, enhancing its ability to generalize across multiple tasks. To validate our proposed framework, we conducted extensive experiments on a variety of datasets, including MTG-Jamendo, DEAM, PMEmo, and EmoMusic. According to our experimental results, the inclusion of musical features, multitask learning, and knowledge distillation significantly enhances performance. In particular, our model outperforms the state-of-the-art models, including the best-performing model from the MediaEval 2021 competition on the MTG-Jamendo dataset. Our work makes a significant contribution to MER by allowing the combination of categorical and dimensional emotion labels in one unified framework, thus enabling training across datasets.","url_abs":"https://arxiv.org/abs/2502.03979v2","url_pdf":"https://arxiv.org/pdf/2502.03979v2.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"towards-unified-music-emotion-recognition","repo_url":"https://github.com/AMAAI-Lab/Music2Emotion","is_official":1,"mentioned_in_paper":0,"mentioned_in_github":1,"framework":"pytorch","reach":{"status":"ok","spdx":"MIT"}}],"tasks":[{"task_slug":"emotion-recognition","task_name":"Emotion Recognition"},{"task_slug":"knowledge-distillation","task_name":"Knowledge Distillation"},{"task_slug":"music-emotion-recognition","task_name":"Music Emotion Recognition"}],"methods":[{"method_slug":"knowledge-distillation","method_name":"Knowledge Distillation"}],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"syntology_url":"https://syntology.ai/paper/2502.03979","atlas_url":"https://app.syntology.ai/?focus=2502.03979","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2502.03979"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-25T09:33:49+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/AMAAI-Lab/Music2Emotion","reach":{"status":"ok","spdx":"MIT"}}],"summary":{"ran":4,"unverified":1},"by_repo_kind":{"official":{"samples":5,"ran":4,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":0,"samples":[{"code_sha256_prefix":"8504ad451d886c64","entry":"get_audio_paths","repo":"AMAAI-Lab/Music2Emotion","repo_kind":"official","path":"utils/mir_eval_modules.py","file_url":"https://github.com/AMAAI-Lab/Music2Emotion/blob/HEAD/utils/mir_eval_modules.py","link_basis":"harvester_set","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":true,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"8504ad451d886c64"}},{"code_sha256_prefix":"561d8d0de9adaf3b","entry":"get_lab_paths","repo":"AMAAI-Lab/Music2Emotion","repo_kind":"official","path":"utils/mir_eval_modules.py","file_url":"https://github.com/AMAAI-Lab/Music2Emotion/blob/HEAD/utils/mir_eval_modules.py","link_basis":"harvester_set","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":true,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"561d8d0de9adaf3b"}},{"code_sha256_prefix":"e9f99ab48d5eb206","entry":"normalize_chord","repo":"AMAAI-Lab/Music2Emotion","repo_kind":"official","path":"music2emo.py","file_url":"https://github.com/AMAAI-Lab/Music2Emotion/blob/HEAD/music2emo.py","link_basis":"harvester_set","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"e9f99ab48d5eb206"}},{"code_sha256_prefix":"1fefbb553a9045c7","entry":"sanitize_key_signature","repo":"AMAAI-Lab/Music2Emotion","repo_kind":"official","path":"music2emo.py","file_url":"https://github.com/AMAAI-Lab/Music2Emotion/blob/HEAD/music2emo.py","link_basis":"harvester_set","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":true,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"1fefbb553a9045c7"}},{"code_sha256_prefix":"f9380e30c46b4699","entry":"gather_all_results","repo":"AMAAI-Lab/Music2Emotion","repo_kind":"official","path":"trainer.py","file_url":"https://github.com/AMAAI-Lab/Music2Emotion/blob/HEAD/trainer.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"f9380e30c46b4699"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}