{"url":"/dataset/fongbespeechdatasetv2","name":"Fongbe Speech Dataset","full_name":"Fongbe speech dataset V2","description_markdown":"This dataset was created for Fongbe automatic speech\r\nrecognition task and contains about 3979 recordings\r\nof 13 participants reading a text written in Fongbe, one\r\nsentence at a time. Fongbe is a vernacular language\r\nspoken mainly in Benin, by more than 50% of the\r\npopulation, and a littke in Togo and in Nigeria. It’s an\r\nunder-resourced because it lacks linguistics resources\r\n(speech corpus and text data) and very few websites\r\nprovide textual data. In this dataset, each example\r\ncontains the audio files and the associated text. The\r\naudio is high-quality (16-bit, 16kHz) recorded using an\r\nadroid app that we built for the need. The dataset is\r\nmulti-speaker, containing recordings from 13\r\nvolunteers (male and female).","description_withheld":null,"homepage":"https://github.com/laleye/FongbeSpeechDataset","introduced_date":"2022-07-01","introduced_date_note":null,"introduced_by":null,"license":null,"modalities":[{"name":"Audio","url":"/datasets/modality/audio"}],"tasks":[{"name":"Automatic Speech Recognition (ASR)","url":"/task/automatic-speech-recognition","datasets_with_task":"/datasets/task/automatic-speech-recognition"}],"languages":[{"name":"Fon","url":"/datasets/language/fon"}],"variants":["Fongbe Speech Dataset"],"data_loaders":[{"repo":"https://github.com/laleye/FongbeSpeechDataset","url":"https://github.com/laleye/FongbeSpeechDataset","frameworks":["pytorch"]}],"num_papers_in_archive":0,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}