{"url":"/method/yellowfin","slug":"yellowfin","name":"YellowFin","full_name":"YellowFin","full_name_withheld":false,"description_markdown":"**YellowFin** is a learning rate and momentum tuner motivated by robustness properties and analysis of quadratic objectives. It stems from a known but obscure fact: the momentum operator's spectral radius is constant in a large subset of the hyperparameter space. For quadratic objectives, the optimizer tunes both the learning rate and the momentum to keep the hyperparameters within a region in which the convergence rate is a constant rate equal to the root momentum. This notion is extended empirically to non-convex objectives. On every iteration, YellowFin optimizes the hyperparameters to minimize a local quadratic optimization.","description_state":"present","introduced_year":null,"introduced_by":{"title":"YellowFin and the Art of Momentum Tuning","paper":"/paper/yellowfin-and-the-art-of-momentum-tuning","first_author":"Jian Zhang","n_authors":2,"url_abs":null,"archive_paper_url":"https://paperswithcode.com/paper/yellowfin-and-the-art-of-momentum-tuning"},"source":{"url":"http://arxiv.org/abs/1706.03471v2","title":"YellowFin and the Art of Momentum Tuning","url_on_a_paper_host":true},"code_snippet_url":"https://github.com/JianGoForIt/YellowFin/blob/847d9948a3707dde501318ba3d31db2c4c7e3e61/tuner_utils/yellowfin.py#L20","code_snippet_url_on_a_code_host":true,"categories":[{"area":"General","area_id":"general","collection":"Stochastic Optimization","url":"/methods/category/stochastic-optimization","pwc_aliases":[]}],"n_papers_tagged":1,"archive_num_papers":1,"papers_newest_first":[{"paper":"/paper/yellowfin-and-the-art-of-momentum-tuning","title":"YellowFin and the Art of Momentum Tuning","date":"2017-06-12","arxiv_id":"1706.03471","n_code_links":2,"syntology":{"ran":0,"of":1,"unverified":1,"pointer_only":0}}],"papers_shown":1,"tasks":[{"task":"/task/constituency-parsing","name":"Constituency Parsing","papers":1},{"task":"/task/language-modeling","name":"Language Modeling","papers":1},{"task":"/task/language-modelling","name":"Language Modelling","papers":1}],"tasks_shown":3,"n_tasks":3,"usage_by_year":[{"year":"2017","papers":1}],"row_source":"methods_table","archive":{"source":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","archive_url":"https://paperswithcode.com/method/yellowfin"},"syntology_read_at":"2026-09-24T18:15:14+00:00"}