Datasets › WANDS

WANDS (Wayfair ANnotation Dataset)

Introduced by Yan Chen et al. in WANDS: Dataset for Product Search Relevance Assessment5 Apr 2022 archive 2025-07-28

The dataset contains:

  • 42,994 candidate products with data comprising product class, title, description, attributes, category hierarchy, average rating, and number of reviews
  • 480 search query strings with predicted product class
  • 233,448 (query string, product) human relevance judgments with labels (exact match, partial match, irrelevant)

The purpose of the dataset is to evaluate retrieval models for product search in the e-commerce domain using expert judgment of whether a product is relevant to a given query. It can be used to benchmark different retrieval against each other. As of its publication in 2022, it was to the best of our knowledge the biggest such public dataset.

The accompanying publication describes in depth the annotation guidelines and process used to collect the dataset. It also includes a measure of the quality of the annotation and experimentally compares the dataset's ability to discriminate the effectiveness of different retrieval models vs other comparable evaluation datasets.

Benchmarks archive 2025-07-28

No leaderboard in the archive resolves to this dataset.

Papers archive 2025-07-28

No paper in the archive has a leaderboard row on this dataset; the archive counts 8 papers for it but never published that list.

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

License archive 2025-07-28

MIT License

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • WANDS

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections