Methods › General › Data Parallel Methods › Local SGD
Local SGD
Introduced by Sebastian U. Stich in Local SGD Converges Fast and Communicates Little
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
Local SGD is a distributed training technique that runs SGD independently in parallel on different workers and averages the sequences only once in a while.
Papers archive 2025-07-28
30 shown of 69, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
DES-LOC: Desynced Low Communication Adaptive Optimizers for Training Foundation Models 28 May 2025 · 0 repositories · arXiv:2505.22549
-
Sharp Gaussian approximations for Decentralized Federated Learning 12 May 2025 · 0 repositories · arXiv:2505.08125
-
Streaming Federated Learning with Markovian Data 24 Mar 2025 · 0 repositories · arXiv:2503.18807
-
EDiT: A Local-SGD-Based Efficient Distributed Training Method for Large Language Models 10 Dec 2024 · 2 repositories · arXiv:2412.07210
-
Collaborative and Efficient Personalization with Mixtures of Adaptors 4 Oct 2024 · 0 repositories · arXiv:2410.03497
-
Does Worst-Performing Agent Lead the Pack? Analyzing Agent Dynamics in Unified Distributed SGD 26 Sep 2024 · 0 repositories · arXiv:2409.17499
-
Convergence of Distributed Adaptive Optimization with Local Updates 20 Sep 2024 · 0 repositories · arXiv:2409.13155
-
Exploring Scaling Laws for Local SGD in Large Language Model Training 20 Sep 2024 · 0 repositories · arXiv:2409.13198
-
Communication-Efficient Adaptive Batch Size Strategies for Distributed Local Gradient Methods 20 Jun 2024 · 0 repositories · arXiv:2406.13936
-
Local Methods with Adaptivity via Scaling 2 Jun 2024 · 0 repositories · arXiv:2406.00846
-
The Limits and Potentials of Local SGD for Distributed Heterogeneous Learning with Intermittent Communication 19 May 2024 · 0 repositories · arXiv:2405.11667
-
Communication Efficient and Provable Federated Unlearning 19 Jan 2024 · 1 repository · arXiv:2401.11018
-
Can We Learn Communication-Efficient Optimizers? 2 Dec 2023 · 1 repository · arXiv:2312.02204
-
Asynchronous SGD on Graphs: a Unified Framework for Asynchronous Decentralized and Federated Optimization 1 Nov 2023 · 0 repositories · arXiv:2311.00465
-
A Quadratic Synchronization Rule for Distributed Deep Learning 22 Oct 2023 · 1 repository · arXiv:2310.14423Syntology ran 5 of 6 samples · 1 unverified
-
Asynchronous Federated Learning with Incentive Mechanism Based on Contract Theory 10 Oct 2023 · 0 repositories · arXiv:2310.06448
-
Stability and Generalization for Minibatch SGD and Local SGD 2 Oct 2023 · 0 repositories · arXiv:2310.01139
-
Global Convergence Analysis of Local SGD for Two-layer Neural Network without Overparameterization 21 Sep 2023 · 0 repositories
-
Preconditioned Federated Learning 20 Sep 2023 · 0 repositories · arXiv:2309.11378
-
FedYolo: Augmenting Federated Learning with Pretrained Transformers 10 Jul 2023 · 0 repositories · arXiv:2307.04905
-
Semi-Asynchronous Federated Edge Learning Mechanism via Over-the-air Computation 6 May 2023 · 0 repositories · arXiv:2305.04066
-
Why (and When) does Local SGD Generalize Better than SGD? 2 Mar 2023 · 1 repository · arXiv:2303.01215Syntology ran 3 of 8 samples · 5 unverified
-
z-SignFedAvg: A Unified Stochastic Sign-based Compression for Federated Learning 6 Feb 2023 · 0 repositories · arXiv:2302.02589
-
When Do Curricula Work in Federated Learning? 24 Dec 2022 · 1 repository · arXiv:2212.12712
-
STSyn: Speeding Up Local SGD with Straggler-Tolerant Synchronization 6 Oct 2022 · 0 repositories · arXiv:2210.03521
-
On the Stability Analysis of Open Federated Learning Systems 25 Sep 2022 · 0 repositories · arXiv:2209.12307
-
FedBR: Improving Federated Learning on Heterogeneous Data via Local Learning Bias Reduction 26 May 2022 · 1 repository · arXiv:2205.13462Syntology ran 9 of 15 samples · 6 unverified
-
Federated Random Reshuffling with Compression and Variance Reduction 8 May 2022 · 0 repositories · arXiv:2205.03914
-
Federated Stochastic Primal-dual Learning with Differential Privacy 26 Apr 2022 · 0 repositories · arXiv:2204.12284
-
ImageNet Challenging Classification with the Raspberry Pi: An Incremental Local Stochastic Gradient Descent Algorithm 21 Mar 2022 · 0 repositories · arXiv:2203.11853
Tasks archive 2025-07-28
20 shown of 23 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections