{"url":"/task/diversity","name":"Diversity","slug":"diversity","description_markdown":"Diversity in data sampling is crucial across various use cases, including search, recommendation systems, and more. Ensuring diverse samples means capturing a wide range of variations and perspectives, which leads to more robust, unbiased, and comprehensive models. In search use cases, for instance, diversity helps avoid redundancy, ensuring that users are exposed to a broader set of relevant information rather than repeated similar results.","categories":[{"name":"Miscellaneous","url":"/area/miscellaneous"}],"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","slug_source":"derived"},"counts":{"papers_tagged":9051,"papers_with_code":3166,"benchmarks":0,"benchmark_tables_in_archive":0,"benchmark_tables_shown":0,"benchmark_tables_withheld_as_spam":0,"benchmark_definition":"a leaderboard table with at least one row; benchmark_tables_shown also counts the zero-row tables; benchmark_tables_in_archive adds the tables withheld as spam","datasets":0,"subtasks":0,"parent_tasks":0},"benchmarks":[],"datasets":[],"subtasks":[],"parent_tasks":[],"papers":{"order":"repositories listed in the archive (desc), then date (desc); the archive holds no stars","population":"papers tagged with this task that list at least one repository in the archive","shown":30,"of":3166,"tagged_in_all":9051,"items":[{"url":"/paper/exploring-the-limits-of-transfer-learning","title":"Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer","date":"2019-10-23","arxiv_id":"1910.10683","repositories_listed":57,"syntology":{"n":31,"n_ran":2,"n_unverified":29,"n_pointer_only":0}},{"url":"/paper/colorful-image-colorization","title":"Colorful Image Colorization","date":"2016-03-28","arxiv_id":"1603.08511","repositories_listed":39,"syntology":{"n":73,"n_ran":32,"n_unverified":41,"n_pointer_only":41}},{"url":"/paper/conditional-image-synthesis-with-auxiliary","title":"Conditional Image Synthesis With Auxiliary Classifier GANs","date":"2016-10-30","arxiv_id":"1610.09585","repositories_listed":37,"syntology":{"n":5,"n_ran":5,"n_unverified":0,"n_pointer_only":5}},{"url":"/paper/diverse-beam-search-decoding-diverse","title":"Diverse Beam Search: Decoding Diverse Solutions from Neural Sequence Models","date":"2016-10-07","arxiv_id":"1610.02424","repositories_listed":25,"syntology":{"n":14,"n_ran":7,"n_unverified":7,"n_pointer_only":14}},{"url":"/paper/the-pile-an-800gb-dataset-of-diverse-text-for","title":"The Pile: An 800GB Dataset of Diverse Text for Language Modeling","date":"2020-12-31","arxiv_id":"2101.00027","repositories_listed":22,"syntology":{"n":1,"n_ran":1,"n_unverified":0,"n_pointer_only":0}},{"url":"/paper/diffusion-models-beat-gans-on-image-synthesis","title":"Diffusion Models Beat GANs on Image Synthesis","date":"2021-05-11","arxiv_id":"2105.05233","repositories_listed":21,"syntology":{"n":50,"n_ran":27,"n_unverified":23,"n_pointer_only":16}},{"url":"/paper/began-boundary-equilibrium-generative","title":"BEGAN: Boundary Equilibrium Generative Adversarial Networks","date":"2017-03-31","arxiv_id":"1703.10717","repositories_listed":18,"syntology":null},{"url":"/paper/the-curious-case-of-neural-text-degeneration","title":"The Curious Case of Neural Text Degeneration","date":"2019-04-22","arxiv_id":"1904.09751","repositories_listed":17,"syntology":{"n":17,"n_ran":8,"n_unverified":9,"n_pointer_only":11}},{"url":"/paper/190600446","title":"Generating Diverse High-Fidelity Images with VQ-VAE-2","date":"2019-06-02","arxiv_id":"1906.00446","repositories_listed":15,"syntology":{"n":9,"n_ran":2,"n_unverified":7,"n_pointer_only":4}},{"url":"/paper/stargan-v2-diverse-image-synthesis-for","title":"StarGAN v2: Diverse Image Synthesis for Multiple Domains","date":"2019-12-04","arxiv_id":"1912.01865","repositories_listed":14,"syntology":{"n":5,"n_ran":0,"n_unverified":5,"n_pointer_only":0}},{"url":"/paper/a-diversity-promoting-objective-function-for","title":"A Diversity-Promoting Objective Function for Neural Conversation Models","date":"2015-10-11","arxiv_id":"1510.03055","repositories_listed":14,"syntology":{"n":13,"n_ran":2,"n_unverified":11,"n_pointer_only":0}},{"url":"/paper/the-ham10000-dataset-a-large-collection-of","title":"The HAM10000 dataset, a large collection of multi-source dermatoscopic images of common pigmented skin lesions","date":"2018-03-28","arxiv_id":"1803.10417","repositories_listed":13,"syntology":{"n":4,"n_ran":0,"n_unverified":4,"n_pointer_only":0}},{"url":"/paper/classifier-free-diffusion-guidance","title":"Classifier-Free Diffusion Guidance","date":"2022-07-26","arxiv_id":"2207.12598","repositories_listed":11,"syntology":{"n":30,"n_ran":21,"n_unverified":9,"n_pointer_only":4}},{"url":"/paper/diffwave-a-versatile-diffusion-model-for","title":"DiffWave: A Versatile Diffusion Model for Audio Synthesis","date":"2020-09-21","arxiv_id":"2009.09761","repositories_listed":11,"syntology":{"n":33,"n_ran":20,"n_unverified":13,"n_pointer_only":0}},{"url":"/paper/a-learned-representation-for-artistic-style","title":"A Learned Representation For Artistic Style","date":"2016-10-24","arxiv_id":"1610.07629","repositories_listed":11,"syntology":null},{"url":"/paper/camera-style-adaptation-for-person-re","title":"Camera Style Adaptation for Person Re-identification","date":"2017-11-28","arxiv_id":"1711.10295","repositories_listed":10,"syntology":{"n":3,"n_ran":0,"n_unverified":3,"n_pointer_only":3}},{"url":"/paper/scalability-in-perception-for-autonomous","title":"Scalability in Perception for Autonomous Driving: Waymo Open Dataset","date":"2019-12-10","arxiv_id":"1912.04838","repositories_listed":9,"syntology":{"n":10,"n_ran":0,"n_unverified":10,"n_pointer_only":0}},{"url":"/paper/wikihow-a-large-scale-text-summarization","title":"WikiHow: A Large Scale Text Summarization Dataset","date":"2018-10-18","arxiv_id":"1810.09305","repositories_listed":9,"syntology":{"n":1,"n_ran":1,"n_unverified":0,"n_pointer_only":1}},{"url":"/paper/unrolled-generative-adversarial-networks","title":"Unrolled Generative Adversarial Networks","date":"2016-11-07","arxiv_id":"1611.02163","repositories_listed":9,"syntology":null},{"url":"/paper/hierarchical-text-conditional-image","title":"Hierarchical Text-Conditional Image Generation with CLIP Latents","date":"2022-04-13","arxiv_id":"2204.06125","repositories_listed":8,"syntology":{"n":38,"n_ran":29,"n_unverified":9,"n_pointer_only":1}},{"url":"/paper/srflow-learning-the-super-resolution-space","title":"SRFlow: Learning the Super-Resolution Space with Normalizing Flow","date":"2020-06-25","arxiv_id":"2006.14200","repositories_listed":8,"syntology":{"n":3,"n_ran":2,"n_unverified":1,"n_pointer_only":3}},{"url":"/paper/deep-reinforcement-learning-for-dialogue","title":"Deep Reinforcement Learning for Dialogue Generation","date":"2016-06-05","arxiv_id":"1606.01541","repositories_listed":8,"syntology":null},{"url":"/paper/underwater-image-enhancement-via-medium","title":"Underwater Image Enhancement via Medium Transmission-Guided Multi-Color Space Embedding","date":"2021-04-27","arxiv_id":"2104.13015","repositories_listed":7,"syntology":{"n":1,"n_ran":1,"n_unverified":0,"n_pointer_only":0}},{"url":"/paper/clotho-an-audio-captioning-dataset","title":"Clotho: An Audio Captioning Dataset","date":"2019-10-21","arxiv_id":"1910.09387","repositories_listed":7,"syntology":{"n":21,"n_ran":6,"n_unverified":15,"n_pointer_only":0}},{"url":"/paper/190410509","title":"Generating Long Sequences with Sparse Transformers","date":"2019-04-23","arxiv_id":"1904.10509","repositories_listed":7,"syntology":{"n":6,"n_ran":5,"n_unverified":1,"n_pointer_only":0}},{"url":"/paper/diverse-image-to-image-translation-via","title":"Diverse Image-to-Image Translation via Disentangled Representations","date":"2018-08-02","arxiv_id":"1808.00948","repositories_listed":7,"syntology":null},{"url":"/paper/pacgan-the-power-of-two-samples-in-generative","title":"PacGAN: The power of two samples in generative adversarial networks","date":"2017-12-12","arxiv_id":"1712.04086","repositories_listed":7,"syntology":null},{"url":"/paper/toward-multimodal-image-to-image-translation","title":"Toward Multimodal Image-to-Image Translation","date":"2017-11-30","arxiv_id":"1711.11586","repositories_listed":7,"syntology":{"n":1,"n_ran":0,"n_unverified":1,"n_pointer_only":1}},{"url":"/paper/using-millions-of-emoji-occurrences-to-learn","title":"Using millions of emoji occurrences to learn any-domain representations for detecting sentiment, emotion and sarcasm","date":"2017-08-01","arxiv_id":"1708.00524","repositories_listed":7,"syntology":{"n":1,"n_ran":1,"n_unverified":0,"n_pointer_only":0}},{"url":"/paper/scaling-language-image-pre-training-via","title":"Scaling Language-Image Pre-training via Masking","date":"2022-12-01","arxiv_id":"2212.00794","repositories_listed":6,"syntology":{"n":14,"n_ran":10,"n_unverified":4,"n_pointer_only":0}}],"syntology_records":24,"syntology_note":"a paper without a record is not a recorded non-run: it may lack an arXiv id or simply be absent from the graph layer"},"description_links":{"kept":0,"unwrapped_to_text":0,"bare_urls_linked":0,"relative_images_dropped":0,"rule":"internal links are kept only when the target slug exists in the catalog"},"syntology":{"read_at":"2026-09-24T18:15:14+00:00","claim":"Per-sample execution status on synthesized fixtures ('ran N of M samples'); not a correctness claim and not a ranking signal.","status_vocabulary":{"ran_honours":"ran, honoured the contract we drafted","ran_violates":"ran, violated the contract we drafted","ran_draft_wrong":"ran; our contract draft was wrong, not the code","ran_fixture":"ran; our fixture could not drive it","ran":"ran on a synthesized input","unverified":"unverified (harvested, no recorded run)"}},"not_shown":{"libraries":"the archive has no per-task library table","trend_sparklines":"the Trend column of the benchmarks table was a rendered image; it is not in the archive","social_and_latest_sorts":"stars and social signals are not in the archive"}}