{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/a-practical-guide-to-statistical-distances","title":"A Practical Guide to Sample-based Statistical Distances for Evaluating Generative Models in Science","arxiv_id":"2403.12636","date":"2024-03-19","proceeding":null,"authors":["Sebastian Bischoff","Alana Darcher","Michael Deistler","Richard Gao","Franziska Gerken","Manuel Gloeckler","Lisa Haxel","Jaivardhan Kapoor","Janne K Lappalainen","Jakob H Macke","Guy Moss","Matthijs Pals","Felix Pei","Rachel Rapp","A Erdem Sağtekin","Cornelius Schröder","Auguste Schulz","Zinovia Stefanidi","Shoji Toyota","Linda Ulmer","Julius Vetter"],"abstract":"Generative models are invaluable in many fields of science because of their ability to capture high-dimensional and complicated distributions, such as photo-realistic images, protein structures, and connectomes. How do we evaluate the samples these models generate? This work aims to provide an accessible entry point to understanding popular sample-based statistical distances, requiring only foundational knowledge in mathematics and statistics. We focus on four commonly used notions of statistical distances representing different methodologies: Using low-dimensional projections (Sliced-Wasserstein; SW), obtaining a distance using classifiers (Classifier Two-Sample Tests; C2ST), using embeddings through kernels (Maximum Mean Discrepancy; MMD), or neural networks (Fr\\'echet Inception Distance; FID). We highlight the intuition behind each distance and explain their merits, scalability, complexity, and pitfalls. To demonstrate how these distances are used in practice, we evaluate generative models from different scientific domains, namely a model of decision-making and a model generating medical images. We showcase that distinct distances can give different results on similar data. Through this guide, we aim to help researchers to use, interpret, and evaluate statistical distances for generative models in science.","url_abs":"https://arxiv.org/abs/2403.12636v2","url_pdf":"https://arxiv.org/pdf/2403.12636v2.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"a-practical-guide-to-statistical-distances","repo_url":"https://github.com/mackelab/labproject","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":1,"framework":"pytorch","reach":null}],"tasks":[{"task_slug":"decision-making","task_name":"Decision Making"}],"methods":[{"method_slug":"focus","method_name":"Focus"}],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2403.12636","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2403.12636"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/mackelab/labproject","reach":null}],"summary":{"ran_honours":3},"by_repo_kind":{"official":{"samples":3,"ran":3,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":0,"samples":[{"code_sha256_prefix":"d7decad58212031e","entry":"linear_kernel","repo":"mackelab/labproject","repo_kind":"official","path":"labproject/metrics/MMD_torch.py","file_url":"https://github.com/mackelab/labproject/blob/HEAD/labproject/metrics/MMD_torch.py","link_basis":"first_harvest_node","language":"python","status":"ran_honours","verification_level":1,"contract_check":"HONOURS","metamorphic_tier":"invariant","behaviour_fingerprint":true,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"d7decad58212031e"}},{"code_sha256_prefix":"fe7e91ac55afdf57","entry":"polynomial_kernel","repo":"mackelab/labproject","repo_kind":"official","path":"labproject/metrics/MMD_torch.py","file_url":"https://github.com/mackelab/labproject/blob/HEAD/labproject/metrics/MMD_torch.py","link_basis":"first_harvest_node","language":"python","status":"ran_honours","verification_level":1,"contract_check":"HONOURS","metamorphic_tier":"invariant","behaviour_fingerprint":true,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"fe7e91ac55afdf57"}},{"code_sha256_prefix":"da72197697b7ff00","entry":"rbf_kernel","repo":"mackelab/labproject","repo_kind":"official","path":"labproject/metrics/MMD_torch.py","file_url":"https://github.com/mackelab/labproject/blob/HEAD/labproject/metrics/MMD_torch.py","link_basis":"first_harvest_node","language":"python","status":"ran_honours","verification_level":1,"contract_check":"HONOURS","metamorphic_tier":"invariant","behaviour_fingerprint":true,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"da72197697b7ff00"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}