{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/generative-auto-bidding-with-value-guided","title":"Generative Auto-Bidding with Value-Guided Explorations","arxiv_id":"2504.14587","date":"2025-04-20","proceeding":null,"authors":["Jingtong Gao","Yewen Li","Shuai Mao","Nan Jiang","Yejing Wang","Qingpeng Cai","Fei Pan","Peng Jiang","Kun Gai","Bo An","Xiangyu Zhao"],"abstract":"Auto-bidding, with its strong capability to optimize bidding decisions within dynamic and competitive online environments, has become a pivotal strategy for advertising platforms. Existing approaches typically employ rule-based strategies or Reinforcement Learning (RL) techniques. However, rule-based strategies lack the flexibility to adapt to time-varying market conditions, and RL-based methods struggle to capture essential historical dependencies and observations within Markov Decision Process (MDP) frameworks. Furthermore, these approaches often face challenges in ensuring strategy adaptability across diverse advertising objectives. Additionally, as offline training methods are increasingly adopted to facilitate the deployment and maintenance of stable online strategies, the issues of documented behavioral patterns and behavioral collapse resulting from training on fixed offline datasets become increasingly significant. To address these limitations, this paper introduces a novel offline Generative Auto-bidding framework with Value-Guided Explorations (GAVE). GAVE accommodates various advertising objectives through a score-based Return-To-Go (RTG) module. Moreover, GAVE integrates an action exploration mechanism with an RTG-based evaluation method to explore novel actions while ensuring stability-preserving updates. A learnable value function is also designed to guide the direction of action exploration and mitigate Out-of-Distribution (OOD) problems. Experimental results on two offline datasets and real-world deployments demonstrate that GAVE outperforms state-of-the-art baselines in both offline evaluations and online A/B tests. By applying the core methods of this framework, we proudly secured first place in the NeurIPS 2024 competition, 'AIGB Track: Learning Auto-Bidding Agents with Generative Models'.","url_abs":"https://arxiv.org/abs/2504.14587v2","url_pdf":"https://arxiv.org/pdf/2504.14587v2.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"generative-auto-bidding-with-value-guided","repo_url":"https://github.com/applied-machine-learning-lab/gave","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":0,"framework":"pytorch","reach":null}],"tasks":[{"task_slug":"reinforcement-learning-1","task_name":"Reinforcement Learning (RL)"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2504.14587","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2504.14587"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/applied-machine-learning-lab/gave","reach":null}],"summary":{"ran":2,"ran_fixture":1,"unverified":1},"by_repo_kind":{"official":{"samples":4,"ran":3,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":4,"samples":[{"code_sha256_prefix":"3203b47e7dbd1fd8","entry":"Block","repo":"applied-machine-learning-lab/gave","repo_kind":"official","path":"code/bidding_train_env/baseline/dt/dt.py","file_url":"https://github.com/applied-machine-learning-lab/gave/blob/HEAD/code/bidding_train_env/baseline/dt/dt.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"3203b47e7dbd1fd8"}},{"code_sha256_prefix":"e1d87efc6af581e3","entry":"CausalSelfAttention","repo":"applied-machine-learning-lab/gave","repo_kind":"official","path":"code/bidding_train_env/baseline/dt/dt.py","file_url":"https://github.com/applied-machine-learning-lab/gave/blob/HEAD/code/bidding_train_env/baseline/dt/dt.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"e1d87efc6af581e3"}},{"code_sha256_prefix":"363afcb8671faa10","entry":"getScore","repo":"Applied-Machine-Learning-Lab/GAVE","repo_kind":"official","path":"code/bidding_train_env/baseline/dt/dt.py","file_url":"https://github.com/Applied-Machine-Learning-Lab/GAVE/blob/HEAD/code/bidding_train_env/baseline/dt/dt.py","link_basis":"first_harvest_node","language":"python","status":"ran_fixture","verification_level":1,"contract_check":"RAISES","metamorphic_tier":"well_formed","behaviour_fingerprint":true,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"363afcb8671faa10"}},{"code_sha256_prefix":"b267cfa19b90f7ff","entry":"GAVE","repo":"applied-machine-learning-lab/gave","repo_kind":"official","path":"code/bidding_train_env/baseline/dt/dt.py","file_url":"https://github.com/applied-machine-learning-lab/gave/blob/HEAD/code/bidding_train_env/baseline/dt/dt.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"b267cfa19b90f7ff"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}