{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/regulatory-dna-sequence-design-with","title":"Regulatory DNA sequence Design with Reinforcement Learning","arxiv_id":"2503.07981","date":"2025-03-11","proceeding":null,"authors":["Zhao Yang","Bing Su","Chuan Cao","Ji-Rong Wen"],"abstract":"Cis-regulatory elements (CREs), such as promoters and enhancers, are relatively short DNA sequences that directly regulate gene expression. The fitness of CREs, measured by their ability to modulate gene expression, highly depends on the nucleotide sequences, especially specific motifs known as transcription factor binding sites (TFBSs). Designing high-fitness CREs is crucial for therapeutic and bioengineering applications. Current CRE design methods are limited by two major drawbacks: (1) they typically rely on iterative optimization strategies that modify existing sequences and are prone to local optima, and (2) they lack the guidance of biological prior knowledge in sequence optimization. In this paper, we address these limitations by proposing a generative approach that leverages reinforcement learning (RL) to fine-tune a pre-trained autoregressive (AR) model. Our method incorporates data-driven biological priors by deriving computational inference-based rewards that simulate the addition of activator TFBSs and removal of repressor TFBSs, which are then integrated into the RL process. We evaluate our method on promoter design tasks in two yeast media conditions and enhancer design tasks for three human cell types, demonstrating its ability to generate high-fitness CREs while maintaining sequence diversity. The code is available at https://github.com/yangzhao1230/TACO.","url_abs":"https://arxiv.org/abs/2503.07981v1","url_pdf":"https://arxiv.org/pdf/2503.07981v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"regulatory-dna-sequence-design-with","repo_url":"https://github.com/yangzhao1230/taco","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":0,"framework":"pytorch","reach":{"status":"ok","spdx":"MIT"}}],"tasks":[{"task_slug":"reinforcement-learning","task_name":"Reinforcement Learning"},{"task_slug":"reinforcement-learning-1","task_name":"Reinforcement Learning (RL)"},{"task_slug":"reinforcement-learning-2","task_name":"reinforcement-learning"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2503.07981","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2503.07981"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/yangzhao1230/taco","reach":{"status":"ok","spdx":"MIT"}}],"summary":{"unverified":6},"by_repo_kind":{"official":{"samples":6,"ran":0,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":0,"samples":[{"code_sha256_prefix":"20a43164a23f4fa6","entry":"distance","repo":"yangzhao1230/taco","repo_kind":"official","path":"aggregate_mbo.py","file_url":"https://github.com/yangzhao1230/taco/blob/HEAD/aggregate_mbo.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"20a43164a23f4fa6"}},{"code_sha256_prefix":"41321225ece03237","entry":"diversity","repo":"yangzhao1230/taco","repo_kind":"official","path":"dna_optimizers/mbo_optimizaer.py","file_url":"https://github.com/yangzhao1230/taco/blob/HEAD/dna_optimizers/mbo_optimizaer.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"41321225ece03237"}},{"code_sha256_prefix":"c683faca8dc9564c","entry":"get_meme_and_ppms_path","repo":"yangzhao1230/taco","repo_kind":"official","path":"reinforce_mbo.py","file_url":"https://github.com/yangzhao1230/taco/blob/HEAD/reinforce_mbo.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"c683faca8dc9564c"}},{"code_sha256_prefix":"22381d24d1a2d2d6","entry":"get_model_name_or_path","repo":"yangzhao1230/taco","repo_kind":"official","path":"reinforce_mbo.py","file_url":"https://github.com/yangzhao1230/taco/blob/HEAD/reinforce_mbo.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"22381d24d1a2d2d6"}},{"code_sha256_prefix":"dfa03c551cddcb56","entry":"get_params","repo":"yangzhao1230/taco","repo_kind":"official","path":"dna_optimizers/mbo_optimizaer.py","file_url":"https://github.com/yangzhao1230/taco/blob/HEAD/dna_optimizers/mbo_optimizaer.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"dfa03c551cddcb56"}},{"code_sha256_prefix":"f55d5143285bb3ac","entry":"get_prefix_label","repo":"yangzhao1230/taco","repo_kind":"official","path":"reinforce_mbo.py","file_url":"https://github.com/yangzhao1230/taco/blob/HEAD/reinforce_mbo.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"f55d5143285bb3ac"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}