{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/listening-to-sounds-of-silence-for-speech","title":"Listening to Sounds of Silence for Speech Denoising","arxiv_id":"2010.12013","date":"2020-10-22","proceeding":"NeurIPS 2020 12","authors":["Ruilin Xu","Rundi Wu","Yuko Ishiwaka","Carl Vondrick","Changxi Zheng"],"abstract":"We introduce a deep learning model for speech denoising, a long-standing challenge in audio analysis arising in numerous applications. Our approach is based on a key observation about human speech: there is often a short pause between each sentence or word. In a recorded speech signal, those pauses introduce a series of time periods during which only noise is present. We leverage these incidental silent intervals to learn a model for automatic speech denoising given only mono-channel audio. Detected silent intervals over time expose not just pure noise but its time-varying features, allowing the model to learn noise dynamics and suppress it from the speech signal. Experiments on multiple datasets confirm the pivotal role of silent interval detection for speech denoising, and our method outperforms several state-of-the-art denoising methods, including those that accept only audio input (like ours) and those that denoise based on audiovisual input (and hence require more information). We also show that our method enjoys excellent generalization properties, such as denoising spoken languages not seen during training.","url_abs":"https://arxiv.org/abs/2010.12013v1","url_pdf":"https://arxiv.org/pdf/2010.12013v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"listening-to-sounds-of-silence-for-speech","repo_url":"https://github.com/henryxrl/Listening-to-Sound-of-Silence-for-Speech-Denoising","is_official":1,"mentioned_in_paper":0,"mentioned_in_github":0,"framework":"pytorch","reach":null}],"tasks":[{"task_slug":"denoising","task_name":"Denoising"},{"task_slug":"sentence","task_name":"Sentence"},{"task_slug":"speech-denoising","task_name":"Speech Denoising"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"syntology_url":"https://syntology.ai/paper/2010.12013","atlas_url":"https://app.syntology.ai/?focus=2010.12013","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2010.12013"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-25T09:33:49+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/henryxrl/Listening-to-Sound-of-Silence-for-Speech-Denoising","reach":null}],"summary":{"ran":4,"unverified":2},"by_repo_kind":{"official":{"samples":6,"ran":4,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":6,"samples":[{"code_sha256_prefix":"dec4df67c232d10c","entry":"ContextAggNet","repo":"henryxrl/Listening-to-Sound-of-Silence-for-Speech-Denoising","repo_kind":"official","path":"model_2_audio_denoising/audio_denoising_model/networks.py","file_url":"https://github.com/henryxrl/Listening-to-Sound-of-Silence-for-Speech-Denoising/blob/HEAD/model_2_audio_denoising/audio_denoising_model/networks.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"dec4df67c232d10c"}},{"code_sha256_prefix":"5d6655c1e096c560","entry":"ConvBlock","repo":"henryxrl/Listening-to-Sound-of-Silence-for-Speech-Denoising","repo_kind":"official","path":"model_2_audio_denoising/audio_denoising_model/networks.py","file_url":"https://github.com/henryxrl/Listening-to-Sound-of-Silence-for-Speech-Denoising/blob/HEAD/model_2_audio_denoising/audio_denoising_model/networks.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":true,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"5d6655c1e096c560"}},{"code_sha256_prefix":"eedac963eca1f715","entry":"DownConvBlock","repo":"henryxrl/Listening-to-Sound-of-Silence-for-Speech-Denoising","repo_kind":"official","path":"model_2_audio_denoising/audio_denoising_model/networks.py","file_url":"https://github.com/henryxrl/Listening-to-Sound-of-Silence-for-Speech-Denoising/blob/HEAD/model_2_audio_denoising/audio_denoising_model/networks.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":true,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"eedac963eca1f715"}},{"code_sha256_prefix":"abb93377d49ef2a2","entry":"UpConvBlock","repo":"henryxrl/Listening-to-Sound-of-Silence-for-Speech-Denoising","repo_kind":"official","path":"model_2_audio_denoising/audio_denoising_model/networks.py","file_url":"https://github.com/henryxrl/Listening-to-Sound-of-Silence-for-Speech-Denoising/blob/HEAD/model_2_audio_denoising/audio_denoising_model/networks.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":true,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"abb93377d49ef2a2"}},{"code_sha256_prefix":"d37ec58c5d0b7d93","entry":"InpaintNet","repo":"henryxrl/Listening-to-Sound-of-Silence-for-Speech-Denoising","repo_kind":"official","path":"model_2_audio_denoising/audio_denoising_model/networks.py","file_url":"https://github.com/henryxrl/Listening-to-Sound-of-Silence-for-Speech-Denoising/blob/HEAD/model_2_audio_denoising/audio_denoising_model/networks.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"d37ec58c5d0b7d93"}},{"code_sha256_prefix":"6c2ed6d029f15474","entry":"JointModel","repo":"henryxrl/Listening-to-Sound-of-Silence-for-Speech-Denoising","repo_kind":"official","path":"model_2_audio_denoising/audio_denoising_model/networks.py","file_url":"https://github.com/henryxrl/Listening-to-Sound-of-Silence-for-Speech-Denoising/blob/HEAD/model_2_audio_denoising/audio_denoising_model/networks.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"6c2ed6d029f15474"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}