Datasets › Douban

Douban (Douban Conversation Corpus)

Introduced by Yu Wu et al. in Sequential Matching Network: A New Architecture for Multi-turn Response Selection in Retrieval-based Chatbots6 Dec 2016 archive 2025-07-28

We release Douban Conversation Corpus, comprising a training data set, a development set and a test set for retrieval based chatbot. The statistics of Douban Conversation Corpus are shown in the following table.

Train Val Test
session-response pairs 1m 50k 10k
Avg. positive response per session 1 1 1.18
Fless Kappa N\A N\A 0.41
Min turn per session 3 3 3
Max ture per session 98 91 45
Average turn per session 6.69 6.75 5.95
Average Word per utterance 18.56 18.50 20.74

The test data contains 1000 dialogue context, and for each context we create 10 responses as candidates. We recruited three labelers to judge if a candidate is a proper response to the session. A proper response means the response can naturally reply to the message given the context. Each pair received three labels and the majority of the labels was taken as the final decision.


As far as we known, this is the first human-labeled test set for retrieval-based chatbots. The entire corpus link https://www.dropbox.com/s/90t0qtji9ow20ca/DoubanConversaionCorpus.zip?dl=0

Benchmarks archive 2025-07-28

All 4 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

Papers archive 2025-07-28

29 shown of 29 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 81. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Knowledge-aware response selection with semantics underlying multi-turn open-domain conversations 1 2 27 Jul 2023 not harvested
Infinite Recommendation Networks: A Data-Centric Approach 5 1 3 Jun 2022 ran 6 of 13 samples (7 unverified)
A federated graph neural network framework for privacy-preserving personalization 1 1 2 Jun 2022 not harvested
GLocal-K: Global and Local Kernels for Recommender Systems 3 1 27 Aug 2021 not harvested
Uni-Encoder: A Fast and Accurate Response Selection Paradigm for Generation-Based Dialogue Systems 1 2 2 Jun 2021 not harvested
Fine-grained Post-training for Improving Retrieval-based Dialogue Systems 1 1 24 May 2021 not harvested
FedGNN: Federated Graph Neural Network for Privacy-Preserving Recommendation 0 1 9 Feb 2021 not harvested
Dialogue Response Selection with Hierarchical Curriculum Learning 1 1 29 Dec 2020 not harvested
Interpretable Recommender System With Heterogeneous Information: A Geometric Deep Learning Perspective 1 1 20 Sep 2020 not harvested
Do Response Selection Models Really Know What's Next? Utterance Manipulation Strategies for Multi-turn Response Selection 1 1 10 Sep 2020 not harvested
Speaker-Aware BERT for Multi-Turn Response Selection in Retrieval-Based Chatbots 2 1 7 Apr 2020 not harvested
Multi-hop Selector Network for Multi-turn Response Selection in Retrieval-based Chatbots 1 1 1 Nov 2019 not harvested
Scalable Probabilistic Matrix Factorization with Graph-Based Priors 1 1 25 Aug 2019 not harvested
An Effective Domain Adaptive Post-Training Method for BERT in Response Selection 1 1 13 Aug 2019 not harvested
One Time of Interaction May Not Be Enough: Go Deep with an Interaction-over-Interaction Network for Response Selection in Dialogues 1 1 1 Jul 2019 not harvested
Inductive Matrix Completion Based on Graph Neural Networks 3 1 26 Apr 2019 ran 1 of 2 samples (1 unverified; 2 pointer-only for licence)
Poly-encoders: Transformer Architectures and Pre-training Strategies for Fast and Accurate Multi-sentence Scoring 7 1 22 Apr 2019 not harvested
Session-based Social Recommendation via Dynamic Graph Attention Networks 2 1 25 Feb 2019 ran 1 of 13 samples (12 unverified)
Learning Topological Representation for Networks via Hierarchical Sampling 1 1 15 Feb 2019 not harvested
Representation Learning for Heterogeneous Information Networks via Embedding Events 1 1 29 Jan 2019 not harvested
Interactive Matching Network for Multi-Turn Response Selection in Retrieval-Based Chatbots 1 1 7 Jan 2019 not harvested
Multi-Turn Response Selection for Chatbots with Deep Attention Matching Network 1 1 1 Jul 2018 not harvested
Modeling Multi-turn Conversation with Deep Utterance Aggregation 1 1 24 Jun 2018 not harvested
Deep Models of Interactions Across Sets 1 1 7 Mar 2018 not harvested
Graph Convolutional Matrix Completion 17 1 7 Jun 2017 not harvested
Geometric Matrix Completion with Recurrent Multi-Graph Neural Networks 2 1 22 Apr 2017 not harvested
Sequential Matching Network: A New Architecture for Multi-turn Response Selection in Retrieval-based Chatbots 3 1 6 Dec 2016 ran 0 of 3 samples (3 unverified)
Hybrid Recommender System based on Autoencoders 4 2 24 Jun 2016 not harvested
Collaborative Filtering with Graph Information: Consistency and Scalable Methods 2 2 1 Dec 2015 not harvested

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

License archive 2025-07-28

No licence recorded in the archive. Absence here is not a statement about the dataset's terms.

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • Douban
  • Douban Monti

2 variant names, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections