{"url":"/method/mdl","slug":"mdl","name":"MDL","full_name":"Minimum Description Length","full_name_withheld":false,"description_markdown":"**Minimum Description Length** provides a criterion for the selection of models, regardless of their complexity, without the restrictive assumption that the data form a sample from a 'true' distribution.\r\n\r\nExtracted from [scholarpedia](http://scholarpedia.org/article/Minimum_description_length)\r\n\r\n**Source**:\r\n\r\nPaper: [J. Rissanen (1978) Modeling by the shortest data description. Automatica 14, 465-471](https://doi.org/10.1016/0005-1098(78)90005-5)\r\n\r\nBook: [P. D. Grünwald (2007) The Minimum Description Length Principle, MIT Press, June 2007, 570 pages](https://ieeexplore.ieee.org/servlet/opac?bknumber=6267274)","description_state":"present","introduced_year":null,"introduced_by":{"title":null,"paper":null,"first_author":null,"n_authors":0,"url_abs":null,"archive_paper_url":null},"source":{"url":null,"title":null,"url_on_a_paper_host":false},"code_snippet_url":null,"code_snippet_url_on_a_code_host":false,"categories":[{"area":"General","area_id":"general","collection":"AutoML","url":"/methods/category/automl","pwc_aliases":[]}],"n_papers_tagged":97,"archive_num_papers":97,"papers_newest_first":[{"paper":null,"title":"Bridging Predictive Coding and MDL: A Two-Part Code Framework for Deep Learning","date":"2025-05-20","arxiv_id":"2505.14635","n_code_links":0,"syntology":null},{"paper":"/paper/a-minimum-description-length-approach-to","title":"A Minimum Description Length Approach to Regularization in Neural Networks","date":"2025-05-19","arxiv_id":"2505.13398","n_code_links":1,"syntology":{"ran":2,"of":2,"unverified":0,"pointer_only":0}},{"paper":null,"title":"A Theory of Machine Understanding via the Minimum Description Length Principle","date":"2025-04-01","arxiv_id":"2504.00395","n_code_links":0,"syntology":null},{"paper":null,"title":"A Materials Map Integrating Experimental and Computational Data via Graph-Based Machine Learning for Enhanced Materials Discovery","date":"2025-03-10","arxiv_id":"2503.07378","n_code_links":0,"syntology":null},{"paper":null,"title":"Comparative Analysis of MDL-VAE vs. Standard VAE on 202 Years of Gynecological Data","date":"2025-02-25","arxiv_id":"2502.18412","n_code_links":0,"syntology":null},{"paper":null,"title":"On Calibration in Multi-Distribution Learning","date":"2024-12-18","arxiv_id":"2412.14142","n_code_links":0,"syntology":null},{"paper":null,"title":"AdaptiveMDL-GenClust: A Robust Clustering Framework Integrating Normalized Mutual Information and Evolutionary Algorithms","date":"2024-11-26","arxiv_id":"2412.05305","n_code_links":0,"syntology":null},{"paper":null,"title":"Interpretability as Compression: Reconsidering SAE Explanations of Neural Activations with MDL-SAEs","date":"2024-10-15","arxiv_id":"2410.11179","n_code_links":0,"syntology":null},{"paper":null,"title":"Neural Networks Generalize on Low Complexity Data","date":"2024-09-19","arxiv_id":"2409.12446","n_code_links":0,"syntology":null},{"paper":null,"title":"MDL-Pool: Adaptive Multilevel Graph Pooling Based on Minimum Description Length","date":"2024-09-16","arxiv_id":"2409.10263","n_code_links":0,"syntology":null},{"paper":null,"title":"A Systematic Review of Intermediate Fusion in Multimodal Deep Learning for Biomedical Applications","date":"2024-08-02","arxiv_id":"2408.02686","n_code_links":0,"syntology":null},{"paper":null,"title":"Coding for Intelligence from the Perspective of Category","date":"2024-07-01","arxiv_id":"2407.01017","n_code_links":0,"syntology":null},{"paper":null,"title":"Crocodile: Cross Experts Covariance for Disentangled Learning in Multi-Domain Recommendation","date":"2024-05-21","arxiv_id":"2405.12706","n_code_links":0,"syntology":null},{"paper":"/paper/probabilistic-truly-unordered-rule-sets","title":"Probabilistic Truly Unordered Rule Sets","date":"2024-01-18","arxiv_id":"2401.09918","n_code_links":1,"syntology":null},{"paper":null,"title":"Skin cancer diagnosis using NIR spectroscopy data of skin lesions in vivo using machine learning algorithms","date":"2024-01-02","arxiv_id":"2401.01200","n_code_links":0,"syntology":null},{"paper":null,"title":"Distribution-Dependent Rates for Multi-Distribution Learning","date":"2023-12-20","arxiv_id":"2312.13130","n_code_links":0,"syntology":null},{"paper":null,"title":"Optimal Multi-Distribution Learning","date":"2023-12-08","arxiv_id":"2312.05134","n_code_links":0,"syntology":null},{"paper":"/paper/exploring-the-hierarchical-structure-of-human","title":"Exploring the hierarchical structure of human plans via program generation","date":"2023-11-30","arxiv_id":"2311.18644","n_code_links":2,"syntology":null},{"paper":null,"title":"Bridging Algorithmic Information Theory and Machine Learning: A New Approach to Kernel Learning","date":"2023-11-21","arxiv_id":"2311.12624","n_code_links":0,"syntology":null},{"paper":null,"title":"Improved MDL Estimators Using Fiber Bundle of Local Exponential Families for Non-exponential Families","date":"2023-11-07","arxiv_id":"2311.03852","n_code_links":0,"syntology":null},{"paper":null,"title":"KG-MDL: Mining Graph Patterns in Knowledge Graphs with the MDL Principle","date":"2023-09-22","arxiv_id":"2309.12908","n_code_links":0,"syntology":null},{"paper":null,"title":"Decoupled Training: Return of Frustratingly Easy Multi-Domain Learning","date":"2023-09-19","arxiv_id":"2309.10302","n_code_links":0,"syntology":null},{"paper":null,"title":"Variational Density Propagation Continual Learning","date":"2023-08-22","arxiv_id":"2308.11801","n_code_links":0,"syntology":null},{"paper":null,"title":"A scoping review on multimodal deep learning in biomedical images and texts","date":"2023-07-14","arxiv_id":"2307.07362","n_code_links":0,"syntology":null},{"paper":"/paper/generating-parametric-brdfs-from-natural","title":"Generating Parametric BRDFs from Natural Language Descriptions","date":"2023-06-19","arxiv_id":"2306.15679","n_code_links":1,"syntology":null},{"paper":null,"title":"Perturbation-Based Two-Stage Multi-Domain Active Learning","date":"2023-06-19","arxiv_id":"2306.10700","n_code_links":0,"syntology":null},{"paper":"/paper/the-whole-is-greater-than-the-sum-of-its-3","title":"The Whole Is Greater than the Sum of Its Parts: Improving Music Source Separation by Bridging Network","date":"2023-05-13","arxiv_id":"2305.07855","n_code_links":1,"syntology":null},{"paper":null,"title":"Multi-Domain Learning From Insufficient Annotations","date":"2023-05-04","arxiv_id":"2305.02757","n_code_links":0,"syntology":null},{"paper":null,"title":"A Comprehensive and Versatile Multimodal Deep Learning Approach for Predicting Diverse Properties of Advanced Materials","date":"2023-03-29","arxiv_id":"2303.16412","n_code_links":0,"syntology":null},{"paper":null,"title":"Learning the Finer Things: Bayesian Structure Learning at the Instantiation Level","date":"2023-03-08","arxiv_id":"2303.04339","n_code_links":0,"syntology":null}],"papers_shown":30,"tasks":[{"task":"/task/model-selection","name":"Model Selection","papers":7},{"task":"/task/multi-task-learning","name":"Multi-Task Learning","papers":5},{"task":"/task/multimodal-deep-learning","name":"Multimodal Deep Learning","papers":5},{"task":"/task/descriptive","name":"Descriptive","papers":4},{"task":"/task/active-learning","name":"Active Learning","papers":3},{"task":"/task/bayesian-inference","name":"Bayesian Inference","papers":3},{"task":"/task/clustering","name":"Clustering","papers":3},{"task":"/task/deep-learning","name":"Deep Learning","papers":3},{"task":"/task/domain-adaptation","name":"Domain Adaptation","papers":3},{"task":"/task/classification","name":"General Classification","papers":3},{"task":"/task/image-classification","name":"Image Classification","papers":3},{"task":"/task/reinforcement-learning-1","name":"Reinforcement Learning (RL)","papers":3},{"task":"/task/subgroup-discovery","name":"Subgroup Discovery","papers":3},{"task":"/task/image-classification","name":"image-classification","papers":3},{"task":"/task/reinforcement-learning-2","name":"reinforcement-learning","papers":3},{"task":"/task/attribute","name":"Attribute","papers":2},{"task":"/task/binary-classification","name":"Binary Classification","papers":2},{"task":"/task/causal-inference","name":"Causal Inference","papers":2},{"task":"/task/change-detection","name":"Change Detection","papers":2},{"task":"/task/decision-making","name":"Decision Making","papers":2}],"tasks_shown":20,"n_tasks":101,"usage_by_year":[{"year":"2009","papers":1},{"year":"2012","papers":1},{"year":"2013","papers":3},{"year":"2014","papers":2},{"year":"2015","papers":3},{"year":"2016","papers":3},{"year":"2017","papers":4},{"year":"2018","papers":2},{"year":"2019","papers":15},{"year":"2020","papers":15},{"year":"2021","papers":9},{"year":"2022","papers":7},{"year":"2023","papers":17},{"year":"2024","papers":10},{"year":"2025","papers":5}],"row_source":"methods_table","archive":{"source":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","archive_url":"https://paperswithcode.com/method/mdl"},"syntology_read_at":"2026-09-24T18:15:14+00:00"}