{"url":"/dataset/jamendo-corpus","name":"Jamendo Corpus","full_name":null,"description_markdown":"The **Jamendo Corpus** is a voice detection dataset consisting of 93 songs with Creative Commons license from the [Jamendo](http://www.jamendo.com/) free music sharing website. Segments of each song are annotated as “voice” (sung or spoken) or “no-voice”. The songs constitute a total of about 6 hours of music. The files are all from different artists and represent various genres from mainstream commercial music. The Jamendo audio files are coded in stereo Vorbis OGG 44.1kHz with 112KB/s bitrate. The original split contains 61, 16 and 16 songs in training, validation and testing set, respectively.\n\nSource: [Vocal detection in music with support vector machines](https://perso.telecom-paristech.fr/grichard/Publications/Icassp08_ramona.pdf)\nAudio Source: [https://zenodo.org/record/2585988](https://zenodo.org/record/2585988)","description_withheld":null,"homepage":"https://zenodo.org/record/2585988","introduced_date":"2008-01-01","introduced_date_note":null,"introduced_by":{"paper":null,"title":"Vocal detection in music with support vector machines","first_author":null,"url":"https://perso.telecom-paristech.fr/grichard/Publications/Icassp08_ramona.pdf"},"license":{"name":"CC BY 4.0","url":"https://creativecommons.org/licenses/by/4.0/legalcode"},"modalities":[{"name":"Texts","url":"/datasets/modality/texts"},{"name":"Audio","url":"/datasets/modality/audio"}],"tasks":[],"languages":[{"name":"English","url":"/datasets/language/english"},{"name":"French","url":"/datasets/language/french"}],"variants":["Jamendo Corpus"],"data_loaders":[],"num_papers_in_archive":3,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}