Methods › General › Learning Rate Schedules › Step Decay
Step Decay
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
Step Decay is a learning rate schedule that drops the learning rate by a factor every few epochs, where the number of epochs is a hyperparameter.
Image Credit: Suki Lau
Papers archive 2025-07-28
30 shown of 73, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
A Multi-Power Law for Loss Curve Prediction Across Learning Rate Schedules 17 Mar 2025 · 1 repository · arXiv:2503.12811Syntology ran 0 of 9 samples · 9 unverified
-
Regularized Xception for facial expression recognition with extra training data and step decay learning rate 17 Oct 2024 · 1 repository
-
Accelerated Convergence of Stochastic Heavy Ball Method under Anisotropic Gradient Noise 22 Dec 2023 · 0 repositories · arXiv:2312.14567
-
DDGM: Solving inverse problems by Diffusive Denoising of Gradient-based Minimization 11 Jul 2023 · 0 repositories · arXiv:2307.04946
-
Gradient Descent, Stochastic Optimization, and Other Tales 2 May 2022 · 0 repositories · arXiv:2205.00832
-
Eigencurve: Optimal Learning Rate Schedule for SGD on Quadratic Objectives with Skewed Hessian Spectrums 27 Oct 2021 · 1 repository · arXiv:2110.14109
-
REX: Revisiting Budgeted Training with an Improved Schedule 9 Jul 2021 · 1 repository · arXiv:2107.04197
-
On the Convergence of Step Decay Step-Size for Stochastic Optimization 18 Feb 2021 · 0 repositories · arXiv:2102.09393
-
Hit-Detector: Hierarchical Trinity Architecture Search for Object Detection 26 Mar 2020 · 1 repository · arXiv:2003.11818
-
Harmonic Convolutional Networks based on Discrete Cosine Transform 18 Jan 2020 · 1 repository · arXiv:2001.06570Syntology ran 0 of 10 samples · 10 unverified
-
MatrixNets: A New Scale and Aspect Ratio Aware Architecture for Object Detection 9 Jan 2020 · 1 repository · arXiv:2001.03194
-
Bridging the Gap Between Anchor-based and Anchor-free Detection via Adaptive Training Sample Selection 5 Dec 2019 · 13 repositories · arXiv:1912.02424Syntology ran 0 of 4 samples · 4 unverified
-
CSPNet: A New Backbone that can Enhance Learning Capability of CNN 27 Nov 2019 · 123 repositories · arXiv:1911.11929
-
Self-training with Noisy Student improves ImageNet classification 11 Nov 2019 · 13 repositories · arXiv:1911.04252Syntology ran 5 of 24 samples · 19 unverified
-
ECA-Net: Efficient Channel Attention for Deep Convolutional Neural Networks 8 Oct 2019 · 13 repositories · arXiv:1910.03151Syntology ran 1 of 8 samples · 7 unverified
-
FreeAnchor: Learning to Match Anchors for Visual Object Detection 5 Sep 2019 · 4 repositories · arXiv:1909.02466Syntology ran 3 of 3 samples · 0 unverified
-
InstaBoost: Boosting Instance Segmentation via Probability Map Guided Copy-Pasting 21 Aug 2019 · 3 repositories · arXiv:1908.07801
-
SCARLET-NAS: Bridging the Gap between Stability and Scalability in Weight-sharing Neural Architecture Search 16 Aug 2019 · 1 repository · arXiv:1908.06022
-
Compact Global Descriptor for Neural Networks 23 Jul 2019 · 1 repository · arXiv:1907.09665
-
Stochastic algorithms with geometric step decay converge linearly on sharp functions 22 Jul 2019 · 1 repository · arXiv:1907.09547
-
Learning Data Augmentation Strategies for Object Detection 26 Jun 2019 · 6 repositories · arXiv:1906.11172Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)
-
Spatial Group-wise Enhance: Improving Semantic Feature Learning in Convolutional Networks 23 May 2019 · 3 repositories · arXiv:1905.09646Syntology ran 2 of 4 samples · 2 unverified · 4 pointer-only (licence)
-
CutMix: Regularization Strategy to Train Strong Classifiers with Localizable Features 13 May 2019 · 30 repositories · arXiv:1905.04899Syntology ran 17 of 24 samples · 7 unverified · 5 pointer-only (licence)
-
Searching for MobileNetV3 6 May 2019 · 67 repositories · arXiv:1905.02244Syntology ran 58 of 105 samples · 47 unverified · 46 pointer-only (licence)
-
The Step Decay Schedule: A Near Optimal, Geometrically Decaying Learning Rate Procedure For Least Squares 29 Apr 2019 · 1 repository · arXiv:1904.12838Syntology ran 1 of 2 samples · 1 unverified
-
GCNet: Non-local Networks Meet Squeeze-Excitation Networks and Beyond 25 Apr 2019 · 9 repositories · arXiv:1904.11492
-
An Energy and GPU-Computation Efficient Backbone Network for Real-Time Object Detection 22 Apr 2019 · 12 repositories · arXiv:1904.09730
-
Data-Driven Neuron Allocation for Scale Aggregation Networks 20 Apr 2019 · 1 repository · arXiv:1904.09460
-
CenterNet: Keypoint Triplets for Object Detection 17 Apr 2019 · 20 repositories · arXiv:1904.08189Syntology ran 2 of 11 samples · 9 unverified · 2 pointer-only (licence)
-
NAS-FPN: Learning Scalable Feature Pyramid Architecture for Object Detection 16 Apr 2019 · 8 repositories · arXiv:1904.07392
Tasks archive 2025-07-28
20 shown of 84 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections