{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/neural-network-compression-framework-for-fast","title":"Neural Network Compression Framework for fast model inference","arxiv_id":"2002.08679","date":"2020-02-20","proceeding":null,"authors":["Alexander Kozlov","Ivan Lazarevich","Vasily Shamporov","Nikolay Lyalyushkin","Yury Gorbachev"],"abstract":"In this work we present a new framework for neural networks compression with fine-tuning, which we called Neural Network Compression Framework (NNCF). It leverages recent advances of various network compression methods and implements some of them, such as sparsity, quantization, and binarization. These methods allow getting more hardware-friendly models which can be efficiently run on general-purpose hardware computation units (CPU, GPU) or special Deep Learning accelerators. We show that the developed methods can be successfully applied to a wide range of models to accelerate the inference time while keeping the original accuracy. The framework can be used within the training samples, which are supplied with it, or as a standalone package that can be seamlessly integrated into the existing training code with minimal adaptations. Currently, a PyTorch version of NNCF is available as a part of OpenVINO Training Extensions at https://github.com/openvinotoolkit/nncf.","url_abs":"https://arxiv.org/abs/2002.08679v4","url_pdf":"https://arxiv.org/pdf/2002.08679v4.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"neural-network-compression-framework-for-fast","repo_url":"https://github.com/openvinotoolkit/nncf","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":0,"framework":"pytorch","reach":{"status":"ok","spdx":"Apache-2.0"}},{"paper_slug":"neural-network-compression-framework-for-fast","repo_url":"https://github.com/openvinotoolkit/nncf_pytorch","is_official":1,"mentioned_in_paper":0,"mentioned_in_github":0,"framework":"pytorch","reach":{"status":"ok","spdx":"Apache-2.0"}}],"tasks":[{"task_slug":"binarization","task_name":"Binarization"},{"task_slug":null,"task_name":"CPU"},{"task_slug":null,"task_name":"GPU"},{"task_slug":"neural-network-compression","task_name":"Neural Network Compression"},{"task_slug":"quantization","task_name":"Quantization"},{"task_slug":"model","task_name":"model"}],"methods":[{"method_slug":"nncf","method_name":"NNCF"}],"datasets_introduced":[],"methods_introduced":[{"slug":"nncf","name":"NNCF","full_name":"Neural Network Compression Framework"}],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2002.08679","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2002.08679"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/openvinotoolkit/nncf_pytorch","reach":{"status":"ok","spdx":"Apache-2.0"}},{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/openvinotoolkit/nncf","reach":{"status":"ok","spdx":"Apache-2.0"}}],"summary":{"unverified":2},"by_repo_kind":{"official":{"samples":2,"ran":0,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":0,"samples":[{"code_sha256_prefix":"aad730edbb0a232f","entry":"deprecated","repo":"openvinotoolkit/nncf","repo_kind":"official","path":"src/nncf/common/deprecation.py","file_url":"https://github.com/openvinotoolkit/nncf/blob/HEAD/src/nncf/common/deprecation.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"Apache-2.0","inline_ok":true,"mcp_get_code":{"code_sha256":"aad730edbb0a232f"}},{"code_sha256_prefix":"a929dd7a8ba33da0","entry":"matches_any","repo":"openvinotoolkit/nncf","repo_kind":"official","path":"src/nncf/common/scopes.py","file_url":"https://github.com/openvinotoolkit/nncf/blob/HEAD/src/nncf/common/scopes.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"Apache-2.0","inline_ok":true,"mcp_get_code":{"code_sha256":"a929dd7a8ba33da0"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}