{"url":"/dataset/deeply-vocal-characterizer","name":"Deeply vocal characterizer","full_name":null,"description_markdown":"Deeply vocal characterizer is a human nonverbal vocalization dataset. This sample dataset consists of about 0.6 hours(56.7 hours in the full set) of audio(16 kHz, 16-bit, mono) across 16 human nonverbal vocalization classes, including throat-clearing, coughing, laughing, panting, and etc. The audio contents are crowdsourced by the general public of South Korea.\r\n\r\nThe dataset is a subset(approximately 1%) of a much bigger dataset which were recorded under the same circumstances as these open-source datasets. Please contact us(contact@deeplyinc.com) for the full set with the research/commercial license.","description_withheld":null,"homepage":"https://github.com/deeplyinc/Vocal-Characterizer-Dataset","introduced_date":"2021-01-27","introduced_date_note":null,"introduced_by":null,"license":null,"modalities":[{"name":"Audio","url":"/datasets/modality/audio"},{"name":"Speech","url":"/datasets/modality/speech"}],"tasks":[],"languages":[],"variants":["Deeply vocal characterizer"],"data_loaders":[{"repo":"https://github.com/deeplyinc/Vocal-Characterizer-Dataset","url":"https://github.com/deeplyinc/Vocal-Characterizer-Dataset","frameworks":[]}],"num_papers_in_archive":0,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}