{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/190509754","title":"A Perceptual Weighting Filter Loss for DNN Training in Speech Enhancement","arxiv_id":"1905.09754","date":"2019-05-23","proceeding":null,"authors":["Ziyue Zhao","Samy Elshamy","Tim Fingscheidt"],"abstract":"Single-channel speech enhancement with deep neural networks (DNNs) has shown\npromising performance and is thus intensively being studied. In this paper,\ninstead of applying the mean squared error (MSE) as the loss function during\nDNN training for speech enhancement, we design a perceptual weighting filter\nloss motivated by the weighting filter as it is employed in\nanalysis-by-synthesis speech coding, e.g., in code-excited linear prediction\n(CELP). The experimental results show that the proposed simple loss function\nimproves the speech enhancement performance compared to a reference DNN with\nMSE loss in terms of perceptual quality and noise attenuation. The proposed\nloss function can be advantageously applied to an existing DNN-based speech\nenhancement system, without modification of the DNN topology for speech\nenhancement. The source code for the proposed approach is made available.","url_abs":"http://arxiv.org/abs/1905.09754v1","url_pdf":"http://arxiv.org/pdf/1905.09754v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"190509754","repo_url":"https://github.com/ifnspaml/perceptual-weighting-filter-loss","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":1,"framework":"tf","reach":null}],"tasks":[{"task_slug":"speech-enhancement","task_name":"Speech Enhancement"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}