{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/deepfilternet-perceptually-motivated-real","title":"DeepFilterNet: Perceptually Motivated Real-Time Speech Enhancement","arxiv_id":"2305.08227","date":"2023-05-14","proceeding":null,"authors":["Hendrik Schröter","Tobias Rosenkranz","Alberto N. Escalante-B.","Andreas Maier"],"abstract":"Multi-frame algorithms for single-channel speech enhancement are able to take advantage from short-time correlations within the speech signal. Deep Filtering (DF) was proposed to directly estimate a complex filter in frequency domain to take advantage of these correlations. In this work, we present a real-time speech enhancement demo using DeepFilterNet. DeepFilterNet's efficiency is enabled by exploiting domain knowledge of speech production and psychoacoustic perception. Our model is able to match state-of-the-art speech enhancement benchmarks while achieving a real-time-factor of 0.19 on a single threaded notebook CPU. The framework as well as pretrained weights have been published under an open source license.","url_abs":"https://arxiv.org/abs/2305.08227v1","url_pdf":"https://arxiv.org/pdf/2305.08227v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"deepfilternet-perceptually-motivated-real","repo_url":"https://github.com/rikorose/deepfilternet","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":1,"framework":"pytorch","reach":{"status":"ok","spdx":"NOASSERTION"}}],"tasks":[{"task_slug":null,"task_name":"CPU"},{"task_slug":"speech-enhancement","task_name":"Speech Enhancement"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[{"leaderboard":"/sota/speech-enhancement-on-demand","task":"Speech Enhancement","dataset":"VoiceBank + DEMAND","model":"DeepFilterNet3","rank_in_archive_order":23,"of":42,"metrics":{"CBAK":"3.61","COVL":"3.77","CSIG":"4.34","PESQ (wb)":"3.17","STOI":"0.944"},"uses_additional_data":false}],"syntology":{"syntology_url":null,"atlas_url":"https://app.syntology.ai/?focus=2305.08227","mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}