{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/local-learning-with-neuron-groups","title":"Local Learning with Neuron Groups","arxiv_id":"2301.07635","date":"2023-01-18","proceeding":null,"authors":["Adeetya Patel","Michael Eickenberg","Eugene Belilovsky"],"abstract":"Traditional deep network training methods optimize a monolithic objective function jointly for all the components. This can lead to various inefficiencies in terms of potential parallelization. Local learning is an approach to model-parallelism that removes the standard end-to-end learning setup and utilizes local objective functions to permit parallel learning amongst model components in a deep network. Recent works have demonstrated that variants of local learning can lead to efficient training of modern deep networks. However, in terms of how much computation can be distributed, these approaches are typically limited by the number of layers in a network. In this work we propose to study how local learning can be applied at the level of splitting layers or modules into sub-components, adding a notion of width-wise modularity to the existing depth-wise modularity associated with local learning. We investigate local-learning penalties that permit such models to be trained efficiently. Our experiments on the CIFAR-10, CIFAR-100, and Imagenet32 datasets demonstrate that introducing width-level modularity can lead to computational advantages over existing methods based on local learning and opens new opportunities for improved model-parallel distributed training. Code is available at: https://github.com/adeetyapatel12/GN-DGL.","url_abs":"https://arxiv.org/abs/2301.07635v1","url_pdf":"https://arxiv.org/pdf/2301.07635v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"local-learning-with-neuron-groups","repo_url":"https://github.com/adeetyapatel12/gn-dgl","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":1,"framework":"pytorch","reach":{"status":"ok","spdx":"MIT"}}],"tasks":[],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2301.07635","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2301.07635"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/adeetyapatel12/gn-dgl","reach":{"status":"ok","spdx":"MIT"}}],"summary":{"unverified":3},"by_repo_kind":{"official":{"samples":3,"ran":0,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":0,"samples":[{"code_sha256_prefix":"73e3dd93c47e13dd","entry":"conv3x3","repo":"adeetyapatel12/gn-dgl","repo_kind":"official","path":"Multilayer_GN_DGL/networks/resnet.py","file_url":"https://github.com/adeetyapatel12/gn-dgl/blob/HEAD/Multilayer_GN_DGL/networks/resnet.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"73e3dd93c47e13dd"}},{"code_sha256_prefix":"1092d1fb39b37d8f","entry":"div_loss_calc","repo":"adeetyapatel12/gn-dgl","repo_kind":"official","path":"Multilayer_GN_DGL/utils.py","file_url":"https://github.com/adeetyapatel12/gn-dgl/blob/HEAD/Multilayer_GN_DGL/utils.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"1092d1fb39b37d8f"}},{"code_sha256_prefix":"f546876d517a9920","entry":"get_ensemble_logits","repo":"adeetyapatel12/gn-dgl","repo_kind":"official","path":"Multilayer_GN_DGL/utils.py","file_url":"https://github.com/adeetyapatel12/gn-dgl/blob/HEAD/Multilayer_GN_DGL/utils.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"f546876d517a9920"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}