{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/autorl-hyperparameter-landscapes","title":"AutoRL Hyperparameter Landscapes","arxiv_id":"2304.02396","date":"2023-04-05","proceeding":null,"authors":["Aditya Mohan","Carolin Benjamins","Konrad Wienecke","Alexander Dockhorn","Marius Lindauer"],"abstract":"Although Reinforcement Learning (RL) has shown to be capable of producing impressive results, its use is limited by the impact of its hyperparameters on performance. This often makes it difficult to achieve good results in practice. Automated RL (AutoRL) addresses this difficulty, yet little is known about the dynamics of the hyperparameter landscapes that hyperparameter optimization (HPO) methods traverse in search of optimal configurations. In view of existing AutoRL approaches dynamically adjusting hyperparameter configurations, we propose an approach to build and analyze these hyperparameter landscapes not just for one point in time but at multiple points in time throughout training. Addressing an important open question on the legitimacy of such dynamic AutoRL approaches, we provide thorough empirical evidence that the hyperparameter landscapes strongly vary over time across representative algorithms from RL literature (DQN, PPO, and SAC) in different kinds of environments (Cartpole, Bipedal Walker, and Hopper) This supports the theory that hyperparameters should be dynamically adjusted during training and shows the potential for more insights on AutoRL problems that can be gained through landscape analyses. Our code can be found at https://github.com/automl/AutoRL-Landscape","url_abs":"https://arxiv.org/abs/2304.02396v4","url_pdf":"https://arxiv.org/pdf/2304.02396v4.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"autorl-hyperparameter-landscapes","repo_url":"https://github.com/automl/autorl-landscape","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":0,"framework":"none","reach":{"status":"ok","spdx":"Apache-2.0"}}],"tasks":[{"task_slug":"automl","task_name":"AutoML"},{"task_slug":"hyperparameter-optimization","task_name":"Hyperparameter Optimization"},{"task_slug":"open-question","task_name":"Open-Ended Question Answering"},{"task_slug":"reinforcement-learning-1","task_name":"Reinforcement Learning (RL)"}],"methods":[{"method_slug":"entropy-regularization","method_name":"Entropy Regularization"},{"method_slug":"ppo","method_name":"PPO"}],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2304.02396","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2304.02396"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/automl/autorl-landscape","reach":{"status":"ok","spdx":"Apache-2.0"}}],"summary":{"unverified":7},"by_repo_kind":{"official":{"samples":7,"ran":0,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":0,"samples":[{"code_sha256_prefix":"3cf934194db5be17","entry":"belongs_to_simplex","repo":"automl/autorl-landscape","repo_kind":"official","path":"autorl_landscape/analyze/rubber_band.py","file_url":"https://github.com/automl/autorl-landscape/blob/HEAD/autorl_landscape/analyze/rubber_band.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"Apache-2.0","inline_ok":true,"mcp_get_code":{"code_sha256":"3cf934194db5be17"}},{"code_sha256_prefix":"22ecd337184a2636","entry":"binary_search","repo":"automl/autorl-landscape","repo_kind":"official","path":"autorl_landscape/analyze/concavity.py","file_url":"https://github.com/automl/autorl-landscape/blob/HEAD/autorl_landscape/analyze/concavity.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"Apache-2.0","inline_ok":true,"mcp_get_code":{"code_sha256":"22ecd337184a2636"}},{"code_sha256_prefix":"996bc4d6ac250deb","entry":"build_truncated_normal","repo":"automl/autorl-landscape","repo_kind":"official","path":"autorl_landscape/distributions.py","file_url":"https://github.com/automl/autorl-landscape/blob/HEAD/autorl_landscape/distributions.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"Apache-2.0","inline_ok":true,"mcp_get_code":{"code_sha256":"996bc4d6ac250deb"}},{"code_sha256_prefix":"616bcea32a7f99cd","entry":"estimate_model_fit","repo":"automl/autorl-landscape","repo_kind":"official","path":"autorl_landscape/ls_models/rbf.py","file_url":"https://github.com/automl/autorl-landscape/blob/HEAD/autorl_landscape/ls_models/rbf.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"Apache-2.0","inline_ok":true,"mcp_get_code":{"code_sha256":"616bcea32a7f99cd"}},{"code_sha256_prefix":"567bb5404403d523","entry":"find_peaks","repo":"automl/autorl-landscape","repo_kind":"official","path":"autorl_landscape/analyze/peaks.py","file_url":"https://github.com/automl/autorl-landscape/blob/HEAD/autorl_landscape/analyze/peaks.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"Apache-2.0","inline_ok":true,"mcp_get_code":{"code_sha256":"567bb5404403d523"}},{"code_sha256_prefix":"e8c362495ced040b","entry":"get_half_of_convex_hull","repo":"automl/autorl-landscape","repo_kind":"official","path":"autorl_landscape/analyze/rubber_band.py","file_url":"https://github.com/automl/autorl-landscape/blob/HEAD/autorl_landscape/analyze/rubber_band.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"Apache-2.0","inline_ok":true,"mcp_get_code":{"code_sha256":"e8c362495ced040b"}},{"code_sha256_prefix":"dafee1bc6703d6cb","entry":"is_picked","repo":"automl/autorl-landscape","repo_kind":"official","path":"autorl_landscape/visualize.py","file_url":"https://github.com/automl/autorl-landscape/blob/HEAD/autorl_landscape/visualize.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"Apache-2.0","inline_ok":true,"mcp_get_code":{"code_sha256":"dafee1bc6703d6cb"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}