Datasets › SNIPS

SNIPS (SNIPS Natural Language Understanding benchmark)

Introduced by Alice Coucke et al. in Snips Voice Platform: an embedded Spoken Language Understanding system for private-by-design voice interfaces1 Jan 2018 archive 2025-07-28

The SNIPS Natural Language Understanding benchmark is a dataset of over 16,000 crowdsourced queries distributed among 7 user intents of various complexity:

  • SearchCreativeWork (e.g. Find me the I, Robot television show),
  • GetWeather (e.g. Is it windy in Boston, MA right now?),
  • BookRestaurant (e.g. I want to book a highly rated restaurant in Paris tomorrow night),
  • PlayMusic (e.g. Play the last track from Beyoncé off Spotify),
  • AddToPlaylist (e.g. Add Diamonds to my roadtrip playlist),
  • RateBook (e.g. Give 6 stars to Of Mice and Men),
  • SearchScreeningEvent (e.g. Check the showtimes for Wonder Woman in Paris). The training set contains of 13,084 utterances, the validation set and the test set contain 700 utterances each, with 100 queries per intent.

Source: https://paperswithcode.com/paper/snips-voice-platform-an-embedded-spoken/

Benchmarks archive 2025-07-28

All 6 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

Papers archive 2025-07-28

14 shown of 14 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 256. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Decomposed Meta-Learning for Few-Shot Sequence Labeling 1 1 4 Mar 2024 not harvested
CTRAN: CNN-Transformer-based Network for Natural Language Understanding 1 2 19 Mar 2023 not harvested
A Hybrid Architecture for Out of Domain Intent Detection and Intent Discovery 1 2 7 Mar 2023 not harvested
CAE: Mechanism to Diminish the Class Imbalanced in SLU Slot Filling Task 1 2 21 Sep 2022 not harvested
Intent Detection and Discovery from User Logs via Deep Semi-Supervised Contrastive Clustering 0 1 1 Jul 2022 not harvested
LIDSNet: A Lightweight on-device Intent Detection model using Deep Siamese Network 0 1 6 Oct 2021 not harvested
Zero-Shot Learning with Common Sense Knowledge Graphs 3 1 18 Jun 2020 not harvested
AGIF: An Adaptive Graph-Interactive Framework for Joint Multiple Intent Detection and Slot Filling 1 2 21 Apr 2020 not harvested
Discovering New Intents via Constrained Deep Adaptive Clustering with Cluster Refinement 1 1 20 Nov 2019 not harvested
Improving Slot Filling by Utilizing Contextual Information 0 1 5 Nov 2019 not harvested
A Stack-Propagation Framework with Token-Level Intent Detection for Spoken Language Understanding 2 4 5 Sep 2019 not harvested
A Novel Bi-directional Interrelated Model for Joint Intent Detection and Slot Filling 2 3 30 Jun 2019 not harvested
Joint Slot Filling and Intent Detection via Capsule Neural Networks 3 2 22 Dec 2018 not harvested
Slot-Gated Modeling for Joint Slot Filling and Intent Prediction 2 2 1 Jun 2018 not harvested

Dataset loaders archive 2025-07-28

3 loaders as listed in the archive; links are outbound and not re-checked here.

Tasks archive 2025-07-28

License archive 2025-07-28

Unknown

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • SNIPS

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections