{"url":"/dataset/voice","name":"VOICe","full_name":null,"description_markdown":"**VOICe** is a dataset for the development and evaluation of domain adaptation methods for sound event detection. VOICe consists of mixtures with three different sound events (\"baby crying\", \"glass breaking\", and \"gunshot\"), which are over-imposed over three different categories of acoustic scenes: vehicle, outdoors, and indoors. Moreover, the mixtures are also offered without any background noise.\r\n\r\nVOICe consists of 1,449 different mixtures of three different sound events:\r\n\r\n* 1,242 mixtures with background noise of three different categories of acoustic scenes (\"vehicle\",\" outdoors\", and \"indoors\"), mixed under 2 SNR values (-3, -9 dB), that is 207 mixtures x 3 acoustic scenes x 2 SNRs = 1,242\r\n* 207 mixtures without any background noise.","description_withheld":null,"homepage":"https://doi.org/10.5281/zenodo.3514950","introduced_date":"2019-11-25","introduced_date_note":null,"introduced_by":{"paper":null,"title":"VOICe: A Sound Event Detection Dataset For Generalizable Domain Adaptation","first_author":null,"url":null},"license":null,"modalities":[{"name":"Audio","url":"/datasets/modality/audio"}],"tasks":[{"name":"Sound Event Detection","url":"/task/sound-event-detection","datasets_with_task":"/datasets/task/sound-event-detection"}],"languages":[],"variants":["VOICe"],"data_loaders":[],"num_papers_in_archive":2,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}