Browse State-of-the-Art › Music Modeling
Music Modeling
24 papers with code · 2 benchmarks · 6 datasets archive 2025-07-28
( Image credit: R-Transformer )
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
2 leaderboard tables shown for this task, 2 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| JSB Chorales (10 rows) | TonicNet | JS Fake Chorales: a Synthetic Dataset of Polyphonic Music with... | code | — | Compare |
| Nottingham (8 rows) | R-Transformer | R-Transformer: Recurrent Neural Network Enhanced Transformer | code | Syntology ran 2 of 2 samples · 0 unverified | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
6 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
24 shown of 24 papers with code (34 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
4 Mar 2018 35 repositories listed Syntology ran 2 of 10 samples · 8 unverified · 1 pointer-only (licence)Our results indicate that a simple convolutional architecture outperforms canonical recurrent networks such as LSTMs across a diverse range of tasks and datasets, while demonstrating longer effective memory.
-
13 Mar 2015 17 repositories listedSeveral variants of the Long Short-Term Memory (LSTM) architecture for recurrent neural networks have been proposed since its inception in 1995.
-
11 Dec 2014 14 repositories listed Syntology ran 0 of 3 samples · 3 unverifiedIn this paper we compare different types of recurrent units in recurrent neural networks (RNNs).
-
12 Sep 2018 12 repositories listed Syntology ran 2 of 4 samples · 2 unverified · 1 pointer-only (licence)This is impractical for long sequences such as musical compositions since their memory complexity for intermediate relative information is quadratic in the sequence length.
-
1 Feb 2020 7 repositories listedIn contrast with this general approach, this paper shows that Transformers can do even better for music modeling, when we improve the way a musical score is converted into the data fed to a Transformer model.
-
18 Mar 2019 4 repositories listed Syntology ran 0 of 1 samples · 1 unverifiedMachine learning models of music typically break up the task of composition into a chronological process, composing a piece of music in a single pass from beginning to end.
-
29 Oct 2018 4 repositories listedGenerating musical audio directly with neural networks is notoriously difficult because it requires coherently modeling structure at many different timescales.
-
29 Mar 2021 3 repositories listed Syntology ran 18 of 34 samples · 16 unverifiedAn important goal of AutoML is to automate-away the design of neural networks on new tasks in under-explored domains.
-
25 Nov 2019 3 repositories listedWe propose a new STAckable Recurrent cell (STAR) for recurrent neural networks (RNNs), which has fewer parameters than widely used LSTM and GRU while being more robust against vanishing or exploding gradients.
-
15 Jan 2020 2 repositories listedThrough the paper, we show how Gaussian mixtures taking into account music metadata information can be used as an effective prior for the autoencoder latent space, introducing the first Music Adversarial Autoencoder…
-
26 Nov 2019 2 repositories listedWe show that training a neural network to predict a seemingly more complex sequence, with extra features included in the series being modelled, can improve overall model performance significantly.
-
12 Jul 2019 2 repositories listed Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)Recurrent Neural Networks have long been the dominating choice for sequence modeling.
-
15 Jun 2016 2 repositories listedOur goal is to be able to build a generative model from a deep neural network architecture to try to create music that has both harmony and melody and is passable as music composed by humans.
-
10 Dec 2024 1 repository listed Syntology ran 0 of 1 samples · 1 unverifiedIn this paper we introduce the Frechet Music Distance (FMD), a novel evaluation metric for generative symbolic music models, inspired by the Frechet Inception Distance (FID) in computer vision and Frechet Audio Distance…
-
12 Oct 2023 1 repository listedSymbolic music is widely used in various deep learning tasks, including generation, transcription, synthesis, and Music Information Retrieval (MIR).
-
2 Dec 2022 1 repository listed Syntology ran 4 of 4 samples · 0 unverified · 4 pointer-only (licence)These important relative attributes, however, are mostly ignored in existing symbolic music modeling methods with the main reason being the lack of a musically-meaningful embedding space where both the absolute and…
-
8 Jan 2022 1 repository listedThis work demonstrates a simple approach to reduce the computational and memory complexity of a large class of structured models.
-
1 Aug 2021 1 repository listedIn this paper, we propose a new recurrent cell called Residual Recurrent Unit (RRU) which beats traditional cells and does not employ a single gate.
-
18 Aug 2020 1 repository listedTo improve harmony, in this paper, we propose a novel MUlti-track MIDI representation (MuMIDI), which enables simultaneous multi-track generation in a single sequence and explicitly models the dependency of the notes…
-
14 Nov 2019 1 repository listed Syntology ran 0 of 6 samples · 6 unverifiedIn comparison to TCN and Wavenet, our network consistently saves memory and computation time, with speed-ups for training and inference of over 4x in the audio generation experiment in particular, while achieving a…
-
25 May 2019 1 repository listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)Long Short-Term Memory (LSTM) infers the long term dependency through a cell state maintained by the input and the forget gate structures, which models a gate output as a value in [0, 1] through a sigmoid function.
-
18 Apr 2017 1 repository listedIn this paper, we propose a new Recurrent Neural Network (RNN) architecture.
-
24 May 2016 1 repository listedHow can we efficiently propagate uncertainty in a latent state representation with recurrent neural networks?
-
27 Jun 2012 1 repository listedWe investigate the problem of modeling symbolic sequences of polyphonic music in a completely general piano-roll representation.
Syntology lines on 10 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections