{"url":"/method/adaptive-loss","slug":"adaptive-loss","name":"Adaptive Loss","full_name":"Adaptive Robust Loss","full_name_withheld":false,"description_markdown":"The Robust Loss is a generalization of the Cauchy/Lorentzian, Geman-McClure, Welsch/Leclerc, generalized Charbonnier, Charbonnier/pseudo-Huber/L1-L2, and L2 loss functions. By introducing robustness as a continuous parameter, the loss function allows algorithms built around robust loss minimization to be generalized, which improves performance on basic vision tasks such as registration and clustering. Interpreting the loss as the negative log of a univariate density yields a general probability distribution that includes normal and Cauchy distributions as special cases. This probabilistic interpretation enables the training of neural networks in which the robustness of the loss automatically adapts itself during training, which improves performance on learning-based tasks such as generative image synthesis and unsupervised monocular depth estimation, without requiring any manual parameter tuning.","description_state":"present","introduced_year":null,"introduced_by":{"title":null,"paper":null,"first_author":null,"n_authors":0,"url_abs":null,"archive_paper_url":null},"source":{"url":"http://arxiv.org/abs/1701.03077v10","title":"A General and Adaptive Robust Loss Function","url_on_a_paper_host":true},"code_snippet_url":null,"code_snippet_url_on_a_code_host":false,"categories":[{"area":"General","area_id":"general","collection":"Loss Functions","url":"/methods/category/loss-functions","pwc_aliases":[]}],"n_papers_tagged":73,"archive_num_papers":null,"papers_newest_first":[{"paper":null,"title":"Loss-Guided Model Sharing and Local Learning Correction in Decentralized Federated Learning for Crop Disease Classification","date":"2025-05-29","arxiv_id":"2505.23063","n_code_links":0,"syntology":null},{"paper":null,"title":"CALF: A Conditionally Adaptive Loss Function to Mitigate Class-Imbalanced Segmentation","date":"2025-04-06","arxiv_id":"2504.04458","n_code_links":0,"syntology":null},{"paper":null,"title":"$μ$KE: Matryoshka Unstructured Knowledge Editing of Large Language Models","date":"2025-04-01","arxiv_id":"2504.01196","n_code_links":0,"syntology":null},{"paper":null,"title":"MagicInfinite: Generating Infinite Talking Videos with Your Words and Voice","date":"2025-03-07","arxiv_id":"2503.05978","n_code_links":0,"syntology":null},{"paper":null,"title":"Variance-Aware Loss Scheduling for Multimodal Alignment in Low-Data Settings","date":"2025-03-05","arxiv_id":"2503.03202","n_code_links":0,"syntology":null},{"paper":null,"title":"APT: Adaptive Personalized Training for Diffusion Models with Limited Data","date":"2025-01-01","arxiv_id":null,"n_code_links":0,"syntology":null},{"paper":null,"title":"Never Reset Again: A Mathematical Framework for Continual Inference in Recurrent Neural Networks","date":"2024-12-20","arxiv_id":"2412.15983","n_code_links":0,"syntology":null},{"paper":null,"title":"Compositional Zero-Shot Learning with Contextualized Cues and Adaptive Contrastive Training","date":"2024-12-10","arxiv_id":"2412.07161","n_code_links":0,"syntology":null},{"paper":"/paper/feddual-a-dual-strategy-with-adaptive-loss","title":"FedDUAL: A Dual-Strategy with Adaptive Loss and Dynamic Aggregation for Mitigating Data Heterogeneity in Federated Learning","date":"2024-12-05","arxiv_id":"2412.04416","n_code_links":1,"syntology":null},{"paper":"/paper/multi-scale-cascaded-large-model-for-whole","title":"Multi-scale Cascaded Large-Model for Whole-body ROI Segmentation","date":"2024-11-23","arxiv_id":"2411.15526","n_code_links":1,"syntology":null},{"paper":null,"title":"A mixed gas concentration regression prediction method based on RESHA-ALW","date":"2024-11-01","arxiv_id":null,"n_code_links":0,"syntology":null},{"paper":null,"title":"Network scaling and scale-driven loss balancing for intelligent poroelastography","date":"2024-10-27","arxiv_id":"2411.08886","n_code_links":0,"syntology":null},{"paper":null,"title":"Spatial-Temporal Mixture-of-Graph-Experts for Multi-Type Crime Prediction","date":"2024-09-24","arxiv_id":"2409.15764","n_code_links":0,"syntology":null},{"paper":"/paper/a-physics-informed-neural-network-pinn","title":"A Physics Informed Neural Network (PINN) Methodology for Coupled Moving Boundary PDEs","date":"2024-09-17","arxiv_id":"2409.10910","n_code_links":1,"syntology":null},{"paper":null,"title":"Real-time Accident Anticipation for Autonomous Driving Through Monocular Depth-Enhanced 3D Modeling","date":"2024-09-02","arxiv_id":"2409.01256","n_code_links":0,"syntology":null},{"paper":null,"title":"Trust And Balance: Few Trusted Samples Pseudo-Labeling and Temperature Scaled Loss for Effective Source-Free Unsupervised Domain Adaptation","date":"2024-09-01","arxiv_id":"2409.00741","n_code_links":0,"syntology":null},{"paper":null,"title":"Hybrid Classification-Regression Adaptive Loss for Dense Object Detection","date":"2024-08-30","arxiv_id":"2408.17182","n_code_links":0,"syntology":null},{"paper":null,"title":"ALTBI: Constructing Improved Outlier Detection Models via Optimization of Inlier-Memorization Effect","date":"2024-08-19","arxiv_id":"2408.09791","n_code_links":0,"syntology":null},{"paper":"/paper/enhancing-thermal-infrared-tracking-with","title":"Coordinate-Aware Thermal Infrared Tracking Via Natural Language Modeling","date":"2024-07-11","arxiv_id":"2407.08265","n_code_links":1,"syntology":null},{"paper":null,"title":"Modeling Temporal Dependencies within the Target for Long-Term Time Series Forecasting","date":"2024-06-07","arxiv_id":"2406.04777","n_code_links":0,"syntology":null},{"paper":null,"title":"Thyroid ultrasound diagnosis improvement via multi-view self-supervised learning and two-stage pre-training","date":"2024-02-18","arxiv_id":"2402.11497","n_code_links":0,"syntology":null},{"paper":null,"title":"Enriched Physics-informed Neural Networks for Dynamic Poisson-Nernst-Planck Systems","date":"2024-02-01","arxiv_id":"2402.01768","n_code_links":0,"syntology":null},{"paper":"/paper/weakly-supervised-semantic-segmentation-for-1","title":"Weakly Supervised Semantic Segmentation for Driving Scenes","date":"2023-12-21","arxiv_id":"2312.13646","n_code_links":1,"syntology":null},{"paper":null,"title":"UMedNeRF: Uncertainty-aware Single View Volumetric Rendering for Medical Neural Radiance Fields","date":"2023-11-10","arxiv_id":"2311.05836","n_code_links":0,"syntology":null},{"paper":null,"title":"Multi-task Learning for Optical Coherence Tomography Angiography (OCTA) Vessel Segmentation","date":"2023-11-03","arxiv_id":"2311.02266","n_code_links":0,"syntology":null},{"paper":null,"title":"L2T-DLN: Learning to Teach with Dynamic Loss Network","date":"2023-10-30","arxiv_id":"2310.19313","n_code_links":0,"syntology":null},{"paper":null,"title":"Combating Label Noise With A General Surrogate Model For Sample Selection","date":"2023-10-16","arxiv_id":"2310.10463","n_code_links":0,"syntology":null},{"paper":null,"title":"Enhanced LFTSformer: A Novel Long-Term Financial Time Series Prediction Model Using Advanced Feature Engineering and the DS Encoder Informer Architecture","date":"2023-10-03","arxiv_id":"2310.01884","n_code_links":0,"syntology":null},{"paper":null,"title":"FourierLoss: Shape-Aware Loss Function with Fourier Descriptors","date":"2023-09-21","arxiv_id":"2309.12106","n_code_links":0,"syntology":null},{"paper":"/paper/weak-pde-learn-a-weak-form-based-approach-to","title":"Weak-PDE-LEARN: A Weak Form Based Approach to Discovering PDEs From Noisy, Limited Data","date":"2023-09-09","arxiv_id":"2309.04699","n_code_links":1,"syntology":null}],"papers_shown":30,"tasks":[{"task":"/task/segmentation","name":"Segmentation","papers":9},{"task":"/task/semantic-segmentation","name":"Semantic Segmentation","papers":8},{"task":"/task/multi-task-learning","name":"Multi-Task Learning","papers":5},{"task":"/task/classification-1","name":"Classification","papers":4},{"task":"/task/denoising","name":"Denoising","papers":4},{"task":"/task/accident-anticipation","name":"Accident Anticipation","papers":3},{"task":"/task/contrastive-learning","name":"Contrastive Learning","papers":3},{"task":"/task/federated-learning","name":"Federated Learning","papers":3},{"task":"/task/classification","name":"General Classification","papers":3},{"task":"/task/meta-learning","name":"Meta-Learning","papers":3},{"task":"/task/object-detection","name":"Object Detection","papers":3},{"task":"/task/object-detection-1","name":"object-detection","papers":3},{"task":"/task/regression-1","name":"regression","papers":3},{"task":"/task/autonomous-driving","name":"Autonomous Driving","papers":2},{"task":"/task/autonomous-vehicles","name":"Autonomous Vehicles","papers":2},{"task":"/task/computed-tomography-ct","name":"Computed Tomography (CT)","papers":2},{"task":"/task/depth-estimation","name":"Depth Estimation","papers":2},{"task":"/task/domain-adaptation","name":"Domain Adaptation","papers":2},{"task":"/task/few-shot-learning","name":"Few-Shot Learning","papers":2},{"task":null,"name":"GPU","papers":2}],"tasks_shown":20,"n_tasks":104,"usage_by_year":[{"year":"2017","papers":1},{"year":"2018","papers":1},{"year":"2019","papers":6},{"year":"2020","papers":9},{"year":"2021","papers":10},{"year":"2022","papers":7},{"year":"2023","papers":17},{"year":"2024","papers":16},{"year":"2025","papers":6}],"row_source":"embedded","archive":{"source":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","archive_url":"https://paperswithcode.com/method/adaptive-loss"},"syntology_read_at":"2026-09-24T18:15:14+00:00"}