{"url":"/dataset/c-and-z","name":"C&Z","full_name":null,"description_markdown":"One of the first datasets (if not the first) to highlight the importance of bias and diversity in the community, which started a revolution afterwards. Introduced in 2014 as integral part of a thesis of Master of Science [1,2] at Carnegie Mellon and City University of Hong Kong. It was later expanded by adding synthetic images generated by a GAN architecture at ETH Zürich (in HDCGAN by Curtó et al. 2017). Being then not only the pioneer of talking about the importance of balanced datasets for learning and vision but also for being the first GAN augmented dataset of faces. \r\n\r\nThe original description goes as follows:\r\n\r\nA bias-free dataset, containing human faces from different ethnical groups in a wide variety of illumination conditions and image resolutions. C&Z is enhanced with HDCGAN synthetic images, thus being the first GAN augmented dataset of faces.\r\n\r\nDataset: [https://github.com/curto2/c](https://github.com/curto2/c)\r\n\r\nSupplement (with scripts to handle the labels): [https://github.com/curto2/graphics](https://github.com/curto2/graphics)\r\n\r\n[1] [https://www.curto.hk/c/decurto.pdf](https://www.curto.hk/c/decurto.pdf)\r\n\r\n[2] [https://www.zarza.hk/z/dezarza.pdf](https://www.zarza.hk/z/dezarza.pdf)","description_withheld":null,"homepage":"https://github.com/curto2/c","introduced_date":"2017-11-17","introduced_date_note":null,"introduced_by":{"paper":"/paper/high-resolution-deep-convolutional-generative","title":"High-Resolution Deep Convolutional Generative Adversarial Networks","first_author":"Joachim D. Curtó","url":null},"license":null,"modalities":[{"name":"Images","url":"/datasets/modality/images"}],"tasks":[],"languages":[],"variants":["C&Z"],"data_loaders":[{"repo":"https://github.com/curto2/c","url":"https://github.com/curto2/c","frameworks":[]}],"num_papers_in_archive":2,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}