{"url":"/dataset/german-credit-dataset","name":"German Credit Dataset","full_name":null,"description_markdown":"Two datasets are provided. the original dataset, in the form provided by Prof. Hofmann, contains categorical/symbolic attributes and is in the file \"german.data\". \r\n\r\nFor algorithms that need numerical attributes, Strathclyde University produced the file \"german.data-numeric\". This file has been edited and several indicator variables added to make it suitable for algorithms which cannot cope with categorical variables. Several attributes that are ordered categorical (such as attribute 17) have been coded as integer. This was the form used by StatLog. \r\n\r\nThis dataset requires use of a cost matrix:\r\n\r\n|   | Good | Bad |\r\n|---|---|---|\r\n| Good | 0 | 1 |\r\n| Bad | 5 | 0 |\r\n\r\nThe rows represent the actual classification and the columns the predicted classification. \r\n\r\nIt is worse to class a customer as good when they are bad (5), than it is to class a customer as bad when they are good (1).","description_withheld":null,"homepage":"https://archive.ics.uci.edu/ml/datasets/statlog+(german+credit+data)","introduced_date":"2022-06-19","introduced_date_note":null,"introduced_by":null,"license":null,"modalities":[{"name":"Tabular","url":"/datasets/modality/tabular"}],"tasks":[],"languages":[],"variants":["German Credit Dataset"],"data_loaders":[],"num_papers_in_archive":21,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}