{"url":"/dataset/minds14","name":"MINDS-14","full_name":null,"description_markdown":"**MINDS-14** is a dataset designed for the **intent detection task with spoken data**. It encompasses **14 distinct intents** extracted from a commercial system in the **e-banking domain**. These intents are associated with spoken examples in **14 diverse language varieties**. The dataset serves as a valuable resource for training and evaluating intent detection models.\r\n\r\nHere are some key details about the **MINDS-14 dataset**:\r\n\r\n- **Tasks**: The primary task is **Automatic Speech Recognition**, specifically **keyword-spotting**.\r\n- **Languages**: The dataset includes examples in **English**, **French**, **Italian**, and **nine other languages**.\r\n- **Multilinguality**: It is a **multilingual** dataset.\r\n- **Size Categories**: The dataset size falls within the range of **10,000 to 100,000 examples**.\r\n- **Language Creators**: The annotations were created by a combination of **crowdsourced** and **expert-generated** contributors.\r\n- **License**: The dataset is available under the **CC-BY-4.0** license.\r\n\r\nSource: Conversation with Bing, 3/18/2024\r\n(1) PolyAI/minds14 · Datasets at Hugging Face. https://huggingface.co/datasets/PolyAI/minds14.\r\n(2) PolyAI/minds14 at main - Hugging Face. https://huggingface.co/datasets/PolyAI/minds14/tree/main.\r\n(3) minds14.py · PolyAI/minds14 at main - Hugging Face. https://huggingface.co/datasets/PolyAI/minds14/blob/main/minds14.py.","description_withheld":null,"homepage":"https://huggingface.co/datasets/PolyAI/minds14","introduced_date":null,"introduced_date_note":null,"introduced_by":null,"license":null,"modalities":[],"tasks":[],"languages":[],"variants":["MINDS-14"],"data_loaders":[],"num_papers_in_archive":0,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}