Home › Datasets › task › Music Generation

Music Generation datasets

archive 2025-07-28

33 datasets carry the task tag "Music Generation" (the task itself: Music Generation), ordered by the archive's paper count. Page 1 of 1: 33 shown of 33. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Music Generation datasets 1–33 of 33

NSynth is a dataset of one shot instrumental notes, containing 305,979 musical notes with unique pitch, timbre and envelope.
138 papers · 2 benchmarks
The MAESTRO dataset contains over 200 hours of paired audio and MIDI recordings from ten years of International Piano-e-Competition.
118 papers · 1 benchmark
MusicCaps is a dataset composed of 5.5k music-text pairs, with rich text descriptions provided by human experts.
84 papers · 1 benchmark
POP909 is a dataset which contains multiple versions of the piano arrangements of 909 popular songs created by professional musicians.
43 papers · 0 benchmarks
URMP (University of Rochester Multi-Modal Musical Performance)
URMP (University of Rochester Multi-Modal Musical Performance) is a dataset for facilitating audio-visual analysis of musical performances.
38 papers · 2 benchmarks
The Lakh MIDI dataset is a collection of 176,581 unique MIDI files, 45,129 of which have been matched and aligned to entries in the Million Song Dataset.
36 papers · 0 benchmarks
The JSB chorales are a set of short, four-voice pieces of music well-noted for their stylistic homogeneity.
33 papers · 1 benchmark
Music21 is an untrimmed video dataset crawled by keyword query from Youtube.
33 papers · 0 benchmarks
MuseData is an electronic library of orchestral and piano classical music from CCARH.
28 papers · 0 benchmarks
EMOPIA (A Multi-Modal Pop Piano Dataset For Emotion Recognition and Emotion-based Music Generation)
EMOPIA (pronounced ‘yee-mò-pi-uh’) dataset is a shared multi-modal (audio and MIDI) database focusing on perceived emotion in pop piano music, to facilitate research on various tasks related to music emotion.
23 papers · 0 benchmarks
Groove (Groove MIDI Dataset)
The Groove MIDI Dataset (GMD) is composed of 13.6 hours of aligned MIDI and (synthesized) audio of human-performed, tempo-aligned expressive drumming.
16 papers · 2 benchmarks
VGMIDI is a dataset of piano arrangements of video game soundtracks.
15 papers · 0 benchmarks
ASAP (Aligned Scores and Performances)
ASAP is a dataset of 222 digital musical scores aligned with 1068 performances (more than 92 hours) of Western classical piano music.
13 papers · 2 benchmarks
The Lakh Pianoroll Dataset (LPD) is a collection of 174,154 multitrack pianorolls derived from the Lakh MIDI Dataset (LMD).
10 papers · 0 benchmarks
NES-MDB (Nintendo Entertainment System Music Database)
The Nintendo Entertainment System Music Database (NES-MDB) is a dataset intended for building automatic music composition systems for the NES audio synthesizer.
7 papers · 0 benchmarks
The MidiCaps dataset [1] is a large-scale dataset of 168,385 midi music files with descriptive text captions, and a set of extracted musical features.
6 papers · 0 benchmarks
The MusicBench dataset is a music audio-text pair dataset that was designed for text-to-music generation purpose and released along with Mustango text-to-music model.
6 papers · 1 benchmark
The ADL Piano MIDI is a dataset of 11,086 piano pieces from different genres.
5 papers · 0 benchmarks
The Song Describer Dataset (SDD) contains ~1.1k captions for 706 permissively licensed music recordings.
5 papers · 1 benchmark
The Bach Doodle Dataset is composed of 21.6 million harmonizations submitted from the Bach Doodle.
4 papers · 0 benchmarks
ComMU has 11,144 MIDI samples that consist of short note sequences created by professional composers with their corresponding 12 metadata.
4 papers · 0 benchmarks
The FakeMusicCaps dataset contains total of 27605 10 seconds music tracks corresponding to almost 77 hours, generated using 5 different Text-To-Music (TTM) models.
4 papers · 0 benchmarks
dMelodies is dataset of simple 2-bar melodies generated using 9 independent latent factors of variation where each data point represents a unique melody based on the following constraints: - Each melody will correspond to a unique scale…
4 papers · 0 benchmarks
A MIDI dataset of 500 4-part chorales generated by the KSChorus algorithm, annotated with results from hundreds of listening test participants, with 500 further unannotated chorales.
3 papers · 0 benchmarks
AIME (AI Music Evaluation Dataset)
The AIME dataset contains 6,000 audio tracks generated by 12 music generation models in addition to 500 tracks from MTG-Jamendo.
2 papers · 0 benchmarks
Guitar-TECHS (Guitar Tones/Techniques, Excerpts & Chords Dataset)
Guitar-TECHS is a comprehensive dataset featuring a variety of guitar techniques, musical excerpts, chords, and scales.
1 paper · 0 benchmarks
JAZZVAR Dataset (JAZZVAR: A Dataset of Variations found within Solo Piano Performances of Jazz Standards for Music Overpainting)
Jazz pianists often uniquely interpret jazz standards.
1 paper · 0 benchmarks
📊 Dataset Details - Name: JamendoMaxCaps - URL: https://huggingface.co/datasets/amaai-lab/JamendoMaxCaps - Content: 362,238 songs with captions generated by Qwen2-Audio Metadata Fields - genre - speed - variable tags 🎯 Rationale 1.
1 paper · 0 benchmarks
NES-VMDB is a dataset containing 98,940 gameplay videos from 389 NES games, each paired with its original soundtrack in symbolic format (MIDI).
1 paper · 0 benchmarks
Introduction The Niko Chord Progression Dataset is used in AccoMontage2.
1 paper · 0 benchmarks
XMIDI is a comprehensive, large-scale symbolic music dataset that includes accurate emotion and genre labels, consisting of 108,023 MIDI files.
1 paper · 0 benchmarks
YM2413-MDB is an 80s FM video game music dataset with multi-label emotion annotations.
1 paper · 0 benchmarks
abc_cc (ABC CC)
Dataset Summary The dataset used to train and evaluate TunesFormer is collected from two sources: The Session and ABCnotation.com.
1 paper · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.