{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/arxiv-2609-01676","title":"Sim2Signal: Sim-to-Real Benchmarks for Traffic Signal Control","arxiv_id":"2609.01676","date":"2026-09-01","proceeding":null,"authors":["Ferdous Al Rafi","Susrik Mukherjee","Latika Liladhar Dekate","Jennifer Yawa Lavoe","Huaiyuan Yao","Shlok Mohanty","Longchao Da","Xuesong Zhou","Hua Wei"],"abstract":"Reinforcement learning achieves strong traffic signal control performance in simulation, yet policies trained in simulators often fail once deployed in the real world, a failure known as the Sim-to-Real gap. When RL is applied to traffic signal control, this gap arises from several sources: sensing, action execution, traffic dynamics, and the control objective. Their relative impact and the reliability of existing Sim-to-Real mitigation methods remain insufficiently understood, and the field lacks a standard benchmark for systematically measuring the gap and evaluating mitigation methods. We present Sim2Signal, a benchmark that decomposes the Sim-to-Real gap into observation, action, transition, and reward gaps, corresponding to mismatches in the four components of the underlying MDP, and induces each gap in isolation under a shared protocol. We evaluate 18 mitigation methods on 2 base controllers, across 33 gap settings and 10 calibrated networks built from 5 real-world locations. We find that direct transfer consistently degrades performance across all four gap sources, but the severity of the degradation does not predict the effectiveness of mitigation. Instead, mitigation effectiveness depends strongly on the network and gap setting: outside the action gap, a method that helps in one case may fail in another. The most effective methods generally estimate what the gap changes, rather than make the policy insensitive through domain randomization or invariant representations. Our code is available at https://github.com/Red-Pheonix/Sim2RealTSCBenchMark","url_abs":"https://arxiv.org/abs/2609.01676","url_pdf":"https://arxiv.org/pdf/2609.01676","source":{"archive":null,"snapshot":"2025-07-28","note":"not in the Papers with Code archive (frozen at the snapshot)","row_kind":"graph","title_abstract_authors_date":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license)"},"code_links":[],"tasks":[],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2609.01676","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2609.01676"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"mentioned_in_github":null,"is_official":null,"provenance":"deterministic:regex_extraction","mentioned_in_paper":null,"url":"https://github.com/Red-Pheonix/Sim2RealTSCBenchMark","reach":{"status":"ok"}}],"summary":{"unverified":3},"by_repo_kind":{"found_in_text":{"samples":3,"ran":0,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":3,"samples":[{"code_sha256_prefix":"a2b53bba5e0d3a58","entry":"merge_same_phase","repo":"Red-Pheonix/Sim2RealTSCBenchMark","repo_kind":"found_in_text","path":"convert_arrows.py","file_url":"https://github.com/Red-Pheonix/Sim2RealTSCBenchMark/blob/HEAD/convert_arrows.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"a2b53bba5e0d3a58"}},{"code_sha256_prefix":"05ca30591ff59901","entry":"parse_pairs","repo":"Red-Pheonix/Sim2RealTSCBenchMark","repo_kind":"found_in_text","path":"convert_arrows.py","file_url":"https://github.com/Red-Pheonix/Sim2RealTSCBenchMark/blob/HEAD/convert_arrows.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"05ca30591ff59901"}},{"code_sha256_prefix":"9f655b37e0d2e5b7","entry":"to_single_arrows","repo":"Red-Pheonix/Sim2RealTSCBenchMark","repo_kind":"found_in_text","path":"convert_arrows.py","file_url":"https://github.com/Red-Pheonix/Sim2RealTSCBenchMark/blob/HEAD/convert_arrows.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"9f655b37e0d2e5b7"}}]},"arxiv_metadata":{"licence":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license)","fields":["title","abstract","authors","date"],"primary_category":"cs.LG","source":"arxiv_daily_20260903.json"},"syntology_extracted_results":null}