{"url":"/dataset/augmented-wine-quality","name":"Augmented Wine Quality","full_name":null,"description_markdown":"The dataset utilized for this study is the Wine Quality dataset, which comprises 1,599 rows and 11 features related to the chemical properties of wine samples. The goal is to predict the ”quality” of the wine, a target variable that is an ordinal integer value, based on the following 10 features: fixed acidity, volatile acidity, citric acid, residual sugar, chlorides, free sulfur dioxide, density, pH, sulphates, and alcohol.\r\n\r\nTo test the model’s ability to identify and classify internal groupings within the dataset, we performed data augmentation. Specifically, we added Gaussian noise to the original feature values, effectively creating subgroups within the data. This noise was calculated as 10% of each feature’s mean and standard deviation, and was added to the data points to generate a new dataset. The result was a dataset that doubled in size to 3,198 rows, simulating internal group structures without providing explicit feature-based indications of these subgroups","description_withheld":null,"homepage":"","introduced_date":"2024-11-27","introduced_date_note":null,"introduced_by":{"paper":"/paper/dynamic-logistic-ensembles-with-recursive","title":"Dynamic Logistic Ensembles with Recursive Probability and Automatic Subset Splitting for Enhanced Binary Classification","first_author":"Mohammad Zubair Khan","url":null},"license":null,"modalities":[],"tasks":[],"languages":[],"variants":["Augmented Wine Quality"],"data_loaders":[],"num_papers_in_archive":1,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}