Datasets › NSynth
NSynth
NSynth is a dataset of one shot instrumental notes, containing 305,979 musical notes with unique pitch, timbre and envelope. The sounds were collected from 1006 instruments from commercial sample libraries and are annotated based on their source (acoustic, electronic or synthetic), instrument family and sonic qualities. The instrument families used in the annotation are bass, brass, flute, guitar, keyboard, mallet, organ, reed, string, synth lead and vocal. Four second monophonic 16kHz audio snippets were generated (notes) for the instruments.
Source: Data Augmentation for Instrument Classification Robust to Audio Effects Image Source: https://magenta.tensorflow.org/nsynth
Benchmarks archive 2025-07-28
All 2 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.
| First row (archive order) | Paper | Code | ||||
|---|---|---|---|---|---|---|
| Few-Shot Audio Classification | NSynth | Meta-Curvature (CRNN) Top-1 Accuracy(5-Way-1-Shot) 96.47 +-0.19 | MetaAudio: A Few-Shot Audio Classification Benchmark | cheggan/metaaudio-a-few-shot-audio-classification-benchmark | 10 | Compare |
| Instrument Recognition | NSynth | M2D-CLAP Accuracy 80.6 | M2D2: Exploring General-purpose Audio-Language... | nttcslab/m2d +1 | 7 | Compare |
Papers archive 2025-07-28
5 shown of 5 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 138. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.
| Date | Samples run Syntology | |||
|---|---|---|---|---|
| M2D2: Exploring General-purpose Audio-Language Representations Beyond CLAP | 2 | 3 | 28 Mar 2025 | not harvested |
| Masked Latent Prediction and Classification for Self-Supervised Audio Representation Learning | 1 | 1 | 17 Feb 2025 | not harvested |
| MT-SLVR: Multi-Task Self-Supervised Learning for Transformation In(Variant) Representations | 1 | 3 | 29 May 2023 | not harvested |
| EfficientLEAF: A Faster LEarnable Audio Frontend of Questionable Use | 1 | 3 | 12 Jul 2022 | not harvested |
| MetaAudio: A Few-Shot Audio Classification Benchmark | 1 | 7 | 5 Apr 2022 | ran 1 of 1 samples (0 unverified; 1 pointer-only for licence) |
Dataset loaders archive 2025-07-28
2 loaders as listed in the archive; links are outbound and not re-checked here.
Tasks archive 2025-07-28
License archive 2025-07-28
Modalities archive 2025-07-28
Languages archive 2025-07-28
No language tagged.
Variants archive 2025-07-28
- NSynth
1 variant name, as the archive lists them.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections