Browse State-of-the-Art › Music Auto-Tagging
Music Auto-Tagging
15 papers with code · 4 benchmarks · 4 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
4 leaderboard tables shown for this task, 4 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| MagnaTagATune (3 rows) | M2D2 AS+ | M2D2: Exploring General-purpose Audio-Language Representations Beyond CLAP | code | — | Compare |
| MagnaTagATune (clean) (3 rows) | Short-chunk CNN + Res | Evaluation of CNN-based Automatic Music Tagging Models | code | Syntology ran 2 of 9 samples · 7 unverified | Compare |
| Million Song Dataset (2 rows) | CLMR | Contrastive Learning of Musical Representations | code | — | Compare |
| TimeTravel (1 row) | Fellini | TräumerAI: Dreaming Music with StyleGAN | code | — | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
4 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
15 shown of 15 papers with code (22 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
6 Mar 2017 3 repositories listed Syntology ran 2 of 3 samples · 1 unverified · 2 pointer-only (licence)Recently, the end-to-end approach that learns hierarchical representations from raw data using deep convolutional neural networks has been successfully explored in the image, text and speech domains.
-
28 Mar 2025 2 repositories listedIn the second stage, it learns CLAP features using the audio features learned from the LLM-based embeddings.
-
24 Oct 2023 2 repositories listedIn this paper, we study whether music source separation can be used as a pre-training strategy for music representation learning, targeted at music classification tasks.
-
17 Oct 2021 2 repositories listedAlong with the evolution of music technology, a large number of styles, or "subgenres," of Electronic Dance Music(EDM) have emerged in recent years.
-
18 Jul 2018 2 repositories listedRecently deep learning based recommendation systems have been actively explored to solve the cold-start problem using a hybrid approach.
-
28 Oct 2017 2 repositories listedRecent work has shown that the end-to-end approach using convolutional neural network (CNN) is effective in various types of machine learning tasks.
-
22 May 2025 1 repository listedMusic auto-tagging is essential for organizing and discovering music in extensive digital libraries.
-
17 Feb 2025 1 repository listedTherefore, we propose a new method: MAsked latenT Prediction And Classification (MATPAC), which is trained with two pretext tasks solved jointly.
-
28 Nov 2024 1 repository listedCommon ways of adapting music foundation models to downstream tasks are probing and fine-tuning.
-
30 Jun 2023 1 repository listedMusic classification has been one of the most popular tasks in the field of music information retrieval.
-
14 Feb 2023 1 repository listedContrastive learning constitutes an emerging branch of self-supervised learning that leverages large amounts of unlabeled data, by learning a latent space, where pairs of different views of the same sample are…
-
17 Mar 2021 1 repository listedA linear classifier trained on the proposed representations achieves a higher average precision than supervised models on the MagnaTagATune dataset, and performs comparably on the Million Song dataset.
-
9 Feb 2021 1 repository listedThe goal of this paper to generate a visually appealing video that responds to music with a neural network so that each frame of the video reflects the musical characteristics of the corresponding audio clip.
-
6 Mar 2017 1 repository listedSecond, we extract audio features from each layer of the pre-trained convolutional networks separately and aggregate them altogether given a long audio clip.
-
20 Aug 2015 1 repository listedFeature learning and deep learning have drawn great attention in recent years as a way of transforming input data into more effective representations using learning algorithms.
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections