{"url":"/dataset/jdsearch","name":"JDsearch","full_name":null,"description_markdown":"**JDsearch** is a personalized product search dataset comprised of real user queries and diverse user-product interaction types (clicking, adding to cart, following, and purchasing) collected from JD.com, a popular Chinese online shopping platform. More specifically, the authors sample about 170,000 active users on a specific date, then record all their interacted products and issued queries in one year, without removing any tail users and products. This finally results in roughly 12,000,000 products, 9,400,000 real searches, and 26,000,000 user-product interactions.\r\n\r\nSource: [JDsearch: A Personalized Product Search Dataset with Real Queries and Full Interactions](https://arxiv.org/pdf/2305.14810v1.pdf)\r\n\r\nImage Source: [JDsearch: A Personalized Product Search Dataset with Real Queries and Full Interactions](https://arxiv.org/pdf/2305.14810v1.pdf)","description_withheld":null,"homepage":"https://github.com/rucliujn/JDsearch","introduced_date":"2023-05-24","introduced_date_note":null,"introduced_by":{"paper":"/paper/jdsearch-a-personalized-product-search","title":"JDsearch: A Personalized Product Search Dataset with Real Queries and Full Interactions","first_author":"Jiongnan Liu","url":null},"license":{"name":"Apache-2.0 license","url":"https://github.com/rucliujn/JDsearch/blob/main/LICENSE"},"modalities":[{"name":"Texts","url":"/datasets/modality/texts"}],"tasks":[],"languages":[],"variants":["JDsearch"],"data_loaders":[],"num_papers_in_archive":3,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}