{"url":"/dataset/3d-speaker","name":"3D-Speaker","full_name":null,"description_markdown":"**3D-Speaker** is a large-scale speech corpus designed to facilitate the research of speech representation disentanglement. 3DSpeaker contains over 10,000 speakers, each of whom are simultaneously recorded by multiple Devices, locating at different Distances, and some speakers are speaking multiple Dialects. The controlled combinations of multi-dimensional audio data yield a matrix of a diverse blend of speech representations entanglement, thereby motivating intriguing methods to untangle them.","description_withheld":null,"homepage":"https://3dspeaker.github.io/","introduced_date":"2023-06-27","introduced_date_note":null,"introduced_by":{"paper":"/paper/3d-speaker-a-large-scale-multi-device-multi","title":"3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement","first_author":"Siqi Zheng","url":null},"license":{"name":"Apache-2.0 license","url":"https://github.com/alibaba-damo-academy/3D-Speaker/blob/main/LICENSE"},"modalities":[{"name":"Audio","url":"/datasets/modality/audio"}],"tasks":[],"languages":[],"variants":["3D-Speaker"],"data_loaders":[],"num_papers_in_archive":1,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}