{"url":"/dataset/fraud-case-verdicts","name":"Fraud_Case_Verdicts","full_name":"The \"Crime Facts\" of \"Offenses of Fraudulence\" in Judicial Yuan Verdicts Dataset","description_markdown":"The \"Crime Facts\" of \"Offenses of Fraudulence\" in Judicial Yuan Verdicts Dataset\r\n\r\nThis data set is based on the judgments of \"Offenses of Fraudulence\" cases published by the Judicial Yuan. The data range of the dataset is from January 1, 2011, to December 31, 2021. 74,823 pieces of original data (judgments and rulings) were collected. We only took the contents of the \"criminal facts\" field of the judgment. This dataset is divided into three parts. The training dataset has 59,858 verdicts, accounting for about 80% of the original data. The remaining 20% ​​is allocated 10% to the verification (7,482 verdicts) and 10% to the test (7,483 verdicts). \"Criminal facts\" have been Chinese word segmented. If Chinese word segmentation is not needed, please merge it yourself.","description_withheld":null,"homepage":"https://huggingface.co/datasets/jslin09/Fraud_Case_Verdicts","introduced_date":"2023-02-10","introduced_date_note":null,"introduced_by":{"paper":"/paper/legal-documents-drafting-with-fine-tuned-pre","title":"Legal Documents Drafting with Fine-Tuned Pre-Trained Large Language Model","first_author":"Chun-Hsien Lin","url":null},"license":{"name":"Apache License v 2.0","url":"https://huggingface.co/datasets/choosealicense/licenses/blob/main/markdown/apache-2.0.md"},"modalities":[],"tasks":[{"name":"Text Generation","url":"/task/text-generation","datasets_with_task":"/datasets/task/text-generation"}],"languages":[{"name":"Chinese","url":"/datasets/language/chinese"}],"variants":["Fraud_Case_Verdicts"],"data_loaders":[],"num_papers_in_archive":1,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}