{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/arxiv-2601-08808","title":"Multiplex Thinking: Reasoning via Token-wise Branch-and-Merge","arxiv_id":"2601.08808","date":"2026-01-13","proceeding":null,"authors":["Yao Tang","Li Dong","Yaru Hao","Qingxiu Dong","Furu Wei","Jiatao Gu"],"abstract":"Large language models often solve complex reasoning tasks more effectively with Chain-of-Thought (CoT), but at the cost of long, low-bandwidth token sequences. Humans, by contrast, often reason softly by maintaining a distribution over plausible next steps. Motivated by this, we propose Multiplex Thinking, a stochastic soft reasoning mechanism that, at each thinking step, samples K candidate tokens and aggregates their embeddings into a single continuous multiplex token. This preserves the vocabulary embedding prior and the sampling dynamics of standard discrete generation, while inducing a tractable probability distribution over multiplex rollouts. Consequently, multiplex trajectories can be directly optimized with on-policy reinforcement learning (RL). Importantly, Multiplex Thinking is self-adaptive: when the model is confident, the multiplex token is nearly discrete and behaves like standard CoT; when it is uncertain, it compactly represents multiple plausible next steps without increasing sequence length. Across challenging math reasoning benchmarks, Multiplex Thinking consistently outperforms strong discrete CoT and RL baselines from Pass@1 through Pass@1024, while producing shorter sequences. The code and checkpoints are available at https://github.com/GMLR-Penn/Multiplex-Thinking.","url_abs":"https://arxiv.org/abs/2601.08808","url_pdf":"https://arxiv.org/pdf/2601.08808","source":{"archive":null,"snapshot":"2025-07-28","note":"not in the Papers with Code archive (frozen at the snapshot)","row_kind":"graph","title_abstract_authors_date":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license)"},"code_links":[],"tasks":[],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2601.08808","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2601.08808"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"mentioned_in_github":null,"is_official":null,"provenance":"deterministic:regex_extraction","mentioned_in_paper":null,"url":"https://github.com/GMLR-Penn/Multiplex-Thinking","reach":null}],"summary":{"ran":7,"unverified":4},"by_repo_kind":{"found_in_text":{"samples":11,"ran":7,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":0,"samples":[{"code_sha256_prefix":"3a6b175c21ec8b40","entry":"ChoicesDecision","repo":"GMLR-Penn/Multiplex-Thinking","repo_kind":"found_in_text","path":"sglang-0.4.9.post6/sglang/lang/ir.py","file_url":"https://github.com/GMLR-Penn/Multiplex-Thinking/blob/HEAD/sglang-0.4.9.post6/sglang/lang/ir.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"3a6b175c21ec8b40"}},{"code_sha256_prefix":"36d0837a8b42abde","entry":"SglConstantText","repo":"GMLR-Penn/Multiplex-Thinking","repo_kind":"found_in_text","path":"sglang-0.4.9.post6/sglang/lang/ir.py","file_url":"https://github.com/GMLR-Penn/Multiplex-Thinking/blob/HEAD/sglang-0.4.9.post6/sglang/lang/ir.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"36d0837a8b42abde"}},{"code_sha256_prefix":"aa7ecaad2f1775f9","entry":"SglExprList","repo":"GMLR-Penn/Multiplex-Thinking","repo_kind":"found_in_text","path":"sglang-0.4.9.post6/sglang/lang/ir.py","file_url":"https://github.com/GMLR-Penn/Multiplex-Thinking/blob/HEAD/sglang-0.4.9.post6/sglang/lang/ir.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"aa7ecaad2f1775f9"}},{"code_sha256_prefix":"536d7d2f26c755db","entry":"SglFork","repo":"GMLR-Penn/Multiplex-Thinking","repo_kind":"found_in_text","path":"sglang-0.4.9.post6/sglang/lang/ir.py","file_url":"https://github.com/GMLR-Penn/Multiplex-Thinking/blob/HEAD/sglang-0.4.9.post6/sglang/lang/ir.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"536d7d2f26c755db"}},{"code_sha256_prefix":"f4b2d7afcef7ccd8","entry":"SglGetForkItem","repo":"GMLR-Penn/Multiplex-Thinking","repo_kind":"found_in_text","path":"sglang-0.4.9.post6/sglang/lang/ir.py","file_url":"https://github.com/GMLR-Penn/Multiplex-Thinking/blob/HEAD/sglang-0.4.9.post6/sglang/lang/ir.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"f4b2d7afcef7ccd8"}},{"code_sha256_prefix":"bde34f1fd07f14b4","entry":"SglSamplingParams","repo":"GMLR-Penn/Multiplex-Thinking","repo_kind":"found_in_text","path":"sglang-0.4.9.post6/sglang/lang/ir.py","file_url":"https://github.com/GMLR-Penn/Multiplex-Thinking/blob/HEAD/sglang-0.4.9.post6/sglang/lang/ir.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"bde34f1fd07f14b4"}},{"code_sha256_prefix":"d8c4cbf4f54d7dc5","entry":"SglSelect","repo":"GMLR-Penn/Multiplex-Thinking","repo_kind":"found_in_text","path":"sglang-0.4.9.post6/sglang/lang/ir.py","file_url":"https://github.com/GMLR-Penn/Multiplex-Thinking/blob/HEAD/sglang-0.4.9.post6/sglang/lang/ir.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"d8c4cbf4f54d7dc5"}},{"code_sha256_prefix":"98e636ab3e712a9f","entry":"ChoicesSamplingMethod","repo":"GMLR-Penn/Multiplex-Thinking","repo_kind":"found_in_text","path":"sglang-0.4.9.post6/sglang/lang/ir.py","file_url":"https://github.com/GMLR-Penn/Multiplex-Thinking/blob/HEAD/sglang-0.4.9.post6/sglang/lang/ir.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"98e636ab3e712a9f"}},{"code_sha256_prefix":"e7911e2a21c1e071","entry":"SglGen","repo":"GMLR-Penn/Multiplex-Thinking","repo_kind":"found_in_text","path":"sglang-0.4.9.post6/sglang/lang/ir.py","file_url":"https://github.com/GMLR-Penn/Multiplex-Thinking/blob/HEAD/sglang-0.4.9.post6/sglang/lang/ir.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"e7911e2a21c1e071"}},{"code_sha256_prefix":"08107026d1aef694","entry":"SglSeparateReasoning","repo":"GMLR-Penn/Multiplex-Thinking","repo_kind":"found_in_text","path":"sglang-0.4.9.post6/sglang/lang/ir.py","file_url":"https://github.com/GMLR-Penn/Multiplex-Thinking/blob/HEAD/sglang-0.4.9.post6/sglang/lang/ir.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"08107026d1aef694"}},{"code_sha256_prefix":"78bfed2f18cbc009","entry":"SglVariable","repo":"GMLR-Penn/Multiplex-Thinking","repo_kind":"found_in_text","path":"sglang-0.4.9.post6/sglang/lang/ir.py","file_url":"https://github.com/GMLR-Penn/Multiplex-Thinking/blob/HEAD/sglang-0.4.9.post6/sglang/lang/ir.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"78bfed2f18cbc009"}}]},"arxiv_metadata":{"licence":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license)","fields":["title","abstract","authors","date"],"primary_category":"cs.LG","source":"arxiv_2026.jsonl"},"syntology_extracted_results":null}