{"url":"/dataset/diakg","name":"DiaKG","full_name":null,"description_markdown":"**DiaKG** is a high-quality Chinese dataset for Diabetes knowledge graph.\r\n\r\nThe dataset is derived from 41 diabetes guidelines and consensus, which are from authoritative Chinese journals including basic research, clinical research, drug usage, clinical cases, diagnosis and treatment methods, etc. The dataset covers the most extensive field of research content and hotspot in recent years. All the annotators have a medical background, and finally conduct a high-quality diabetes database which contains 22,050 entities and 6,890 relations in total. Based on this dataset, doctors, researchers, and enterprise developers can develop knowledge bases for clinical diagnosis, knowledge graphs, and auxiliary diagnostics to further explore the mysteries of diabetes.","description_withheld":null,"homepage":"https://tianchi.aliyun.com/dataset/dataDetail?dataId=88836","introduced_date":"2021-05-31","introduced_date_note":null,"introduced_by":{"paper":"/paper/diakg-an-annotated-diabetes-dataset-for","title":"DiaKG: an Annotated Diabetes Dataset for Medical Knowledge Graph Construction","first_author":"Dejie Chang","url":null},"license":{"name":"Unknown","url":null},"modalities":[{"name":"Texts","url":"/datasets/modality/texts"}],"tasks":[{"name":"Named Entity Recognition (NER)","url":"/task/named-entity-recognition-ner","datasets_with_task":"/datasets/task/named-entity-recognition-ner"},{"name":"Relation Extraction","url":"/task/relation-extraction","datasets_with_task":"/datasets/task/relation-extraction"}],"languages":[{"name":"Chinese","url":"/datasets/language/chinese"}],"variants":["DiaKG"],"data_loaders":[],"num_papers_in_archive":1,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-25T09:33:49+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}