{"url":"/dataset/mudoco-queryrewrite","name":"MuDoCo_QueryRewrite","full_name":"The MuDoCo dataset with Query Rewrite Annotations","description_markdown":"<Task description: joint learning of coreference resolution and query rewrite>\r\n\r\nGiven an ongoing dialogue between a user and a dialogue assistant, for the user query, the model is required to predict both coreference links between the query and the dialogue context, and the self-contained rewritten user query that is independent to the dialogue context.\r\n\r\n<Dataset>\r\n\r\nThe MuDoCo dataset is a public dataset that contains 7.5k task-oriented multi-turn dialogues across 6 domains (calling, messaging, music, news, reminders, weather). Each dialogue turn is annotated with coreference links (links field). Please refer to the paper of the MuDoCo dataset for more details. Upon on the MuDoCo dataset, we annotate the query rewrite for each utterance, including both user and system turns.  More details are provided in https://github.com/apple/ml-cread.","description_withheld":null,"homepage":"https://github.com/apple/ml-cread","introduced_date":"2021-05-20","introduced_date_note":null,"introduced_by":{"paper":"/paper/cread-combined-resolution-of-ellipses-and","title":"CREAD: Combined Resolution of Ellipses and Anaphora in Dialogues","first_author":"Bo-Hsiang Tseng","url":null},"license":null,"modalities":[{"name":"Texts","url":"/datasets/modality/texts"}],"tasks":[{"name":"Coreference Resolution","url":"/task/coreference-resolution","datasets_with_task":"/datasets/task/coreference-resolution"},{"name":"Sentence ReWriting","url":"/task/sentence-rewriting","datasets_with_task":"/datasets/task/sentence-rewriting"},{"name":"Context Query Reformulation","url":"/task/context-query-reformulation","datasets_with_task":"/datasets/task/context-query-reformulation"}],"languages":[{"name":"English","url":"/datasets/language/english"}],"variants":["MuDoCo_QueryRewrite"],"data_loaders":[],"num_papers_in_archive":1,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}