{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/arxiv-2606-29718","title":"Diagnosing and Mitigating Context Rot in Long-horizon Search","arxiv_id":"2606.29718","date":"2026-06-29","proceeding":null,"authors":["Shijie Xia","Yikun Wang","Zhen Huang","Pengfei Liu"],"abstract":"Extensive context has become the norm as Large Language Models (LLMs) are increasingly deployed in long-horizon search tasks. The concern that increasing context length degrades model capabilities, known as context rot, has become a widely recognized issue for these applications. However, in deep search scenarios, it remains unclear how models actually fail under extensive context, and to what extent existing methods can mitigate such failures. Through a systematic study of four flagship models across three benchmarks, we identify a previously overlooked phenomenon, which we term premature termination: under extensive context, models give up or provide uncertain incorrect answers long before exhausting the context window. By controlling for query difficulty, we show that the premature termination rate is positively correlated with context length. Based on the findings, we revisit methods to mitigate context rot, including context management and parallel sampling. For context management, we analyze seven methods across three categories and show that they are inherently test-time scaling strategies that reduce the premature termination rate to enable more exploration, and we further provide model-dependent principles for method selection. For parallel sampling, we develop a behavior-aware filtering strategy and observe a performance gain of 2.6% to 4.9% across three aggregation methods.","url_abs":"https://arxiv.org/abs/2606.29718","url_pdf":"https://arxiv.org/pdf/2606.29718","source":{"archive":null,"snapshot":"2025-07-28","note":"not in the Papers with Code archive (frozen at the snapshot)","row_kind":"graph","title_abstract_authors_date":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license)"},"code_links":[],"tasks":[],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2606.29718","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2606.29718"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"mentioned_in_github":null,"is_official":null,"provenance":"deterministic:regex_extraction","mentioned_in_paper":null,"url":"https://github.com/GAIR-NLP/ContextRot","reach":null}],"summary":{"ran":4,"unverified":1},"by_repo_kind":{"found_in_text":{"samples":5,"ran":4,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":5,"samples":[{"code_sha256_prefix":"31aab5911faefdf2","entry":"build_default_output_csv_path","repo":"GAIR-NLP/ContextRot","repo_kind":"found_in_text","path":"localsearch/analysis/struggle_score.py","file_url":"https://github.com/GAIR-NLP/ContextRot/blob/HEAD/localsearch/analysis/struggle_score.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":true,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"31aab5911faefdf2"}},{"code_sha256_prefix":"c33be04b0b78e278","entry":"classify_struggle_with_llm","repo":"GAIR-NLP/ContextRot","repo_kind":"found_in_text","path":"localsearch/analysis/struggle_score.py","file_url":"https://github.com/GAIR-NLP/ContextRot/blob/HEAD/localsearch/analysis/struggle_score.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"c33be04b0b78e278"}},{"code_sha256_prefix":"6daa23db7ff01bbb","entry":"split_agent_sessions","repo":"GAIR-NLP/ContextRot","repo_kind":"found_in_text","path":"localsearch/analysis/struggle_score.py","file_url":"https://github.com/GAIR-NLP/ContextRot/blob/HEAD/localsearch/analysis/struggle_score.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":true,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"6daa23db7ff01bbb"}},{"code_sha256_prefix":"7b81dbccc81df535","entry":"to_native","repo":"GAIR-NLP/ContextRot","repo_kind":"found_in_text","path":"localsearch/src/eval_bc.py","file_url":"https://github.com/GAIR-NLP/ContextRot/blob/HEAD/localsearch/src/eval_bc.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"7b81dbccc81df535"}},{"code_sha256_prefix":"660a3badfcab0324","entry":"single_round_statistics","repo":"GAIR-NLP/ContextRot","repo_kind":"found_in_text","path":"websearch/src/evaluate.py","file_url":"https://github.com/GAIR-NLP/ContextRot/blob/HEAD/websearch/src/evaluate.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"660a3badfcab0324"}}]},"arxiv_metadata":{"licence":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license)","fields":["title","abstract","authors","date"],"primary_category":"cs.CL","source":"arxiv_2026.jsonl"},"syntology_extracted_results":null}