{"url":"/dataset/crossref","name":"Crossref","full_name":null,"description_markdown":"**Crossref** is an essential organization in the scholarly publishing domain. It plays a crucial role in facilitating the discovery and linking of scholarly content. Let's delve into the details:\r\n\r\n1. **Datasets Registration**:\r\n   - **Dataset records** within Crossref capture information about one or more database records or collections. When registering datasets, it's important to follow these guidelines:\r\n     - Register a **DOI** (or include a registered DOI) for a parent database. Datasets must be registered as part of a collection.\r\n     - Include all relevant **funding**, **license**, and **relationship metadata**.\r\n     - Specify all **contributors** involved.\r\n     - Provide relevant **dates** (creation, publication, and update dates).\r\n     - Describe the dataset's **format** and provide **citation metadata**.\r\n   - If you're interested in the technical details, you can explore the [Datasets Markup Guide](https://www.crossref.org/documentation/principles-practices/datasets/) for XML and metadata assistance¹.\r\n\r\n2. **Free Public Data File**:\r\n   - Crossref has generously made available a **free data file** containing public elements from its **112.5 million metadata records**. This file, which is approximately **65GB** in size and provided in **JSON format**, can be accessed via [Academic Torrents](https://doi.org/10.13003/83B2GP)².\r\n\r\n3. **Metadata Search**:\r\n   - If you're looking for specific scholarly content, you can explore the [Crossref Metadata Search](https://search.crossref.org/). It allows you to search the metadata of journal articles, books, standards, datasets, and more³.\r\n\r\nIn summary, Crossref's dataset initiatives contribute significantly to the scholarly ecosystem by ensuring proper identification, linking, and accessibility of research data.\r\n\r\nSource: Conversation with Bing, 3/17/2024\r\n(1) Datasets - Crossref. https://www.crossref.org/documentation/principles-practices/datasets/.\r\n(2) Free public data file of 112+ million Crossref records. https://www.crossref.org/blog/free-public-data-file-of-112-million-crossref-records/.\r\n(3) Crossref Metadata Search. https://search.crossref.org/.\r\n(4) undefined. https://doi.org/10.13003/83B2GP.","description_withheld":null,"homepage":"https://www.crossref.org/","introduced_date":null,"introduced_date_note":null,"introduced_by":null,"license":null,"modalities":[],"tasks":[],"languages":[],"variants":["Crossref"],"data_loaders":[{"repo":"https://github.com/aconti6/Datasets_FNSFundedProjects","url":"https://github.com/aconti6/Datasets_FNSFundedProjects","frameworks":[]}],"num_papers_in_archive":0,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}