{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/arxiv-2604-24927","title":"Large Language Models Explore by Latent Distilling","arxiv_id":"2604.24927","date":"2026-04-27","proceeding":"ICML","authors":["Yuanhao Zeng","Ao Lu","Lufei Li","Zheng Zhang","Yexin Li","Kan Ren"],"abstract":"Generating diverse responses is crucial for test-time scaling of large language models (LLMs), yet standard stochastic sampling mostly yields surface-level lexical variation, limiting semantic exploration. In this paper, we propose Exploratory Sampling (ESamp), a decoding approach that explicitly encourages semantic diversity during generation. ESamp is motivated by the well-known observation that neural networks tend to make lower-error predictions on inputs similar to those encountered before, and incur higher prediction error on novel ones. Building on this property, we train a lightweight Distiller at test time to predict deep-layer hidden representations of the LLM from its shallow-layer representations to model the LLM's depth-wise representation transitions. During decoding, the Distiller continuously adapts to the mappings induced by the current generation context. ESamp uses the prediction error as a novelty signal to reweight candidate token extensions conditioned on the current prefix, thereby biasing decoding toward less-explored semantic patterns. ESamp is implemented with an asynchronous training--inference pipeline, with less than 5% worst case overhead (1.2% in the optimized release). Empirical results show that ESamp significantly boosts the Pass@k efficiency of reasoning models, showing superior or comparable performance to strong stochastic and heuristic baselines. Notably, ESamp achieves robust generalization across mathematics, science, and code generation benchmarks and breaks the trade-off between diversity and coherence in creative writing. Our code has released at: https://github.com/LinesHogan/tLLM.","url_abs":"https://arxiv.org/abs/2604.24927","url_pdf":"https://arxiv.org/pdf/2604.24927","source":{"archive":null,"snapshot":"2025-07-28","note":"not in the Papers with Code archive (frozen at the snapshot)","row_kind":"graph","title_abstract_authors_date":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license)"},"code_links":[],"tasks":[],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2604.24927","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2604.24927"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"mentioned_in_github":null,"is_official":null,"provenance":"deterministic:regex_extraction","mentioned_in_paper":null,"url":"https://github.com/LinesHogan/tLLM","reach":null}],"summary":{"ran":4,"unverified":1},"by_repo_kind":{"found_in_text":{"samples":5,"ran":4,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":5,"samples":[{"code_sha256_prefix":"8b3ba6ec15c67575","entry":"candidate_capture_paths","repo":"LinesHogan/tLLM","repo_kind":"found_in_text","path":"tllm/common/path_resolution.py","file_url":"https://github.com/LinesHogan/tLLM/blob/HEAD/tllm/common/path_resolution.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":true,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"8b3ba6ec15c67575"}},{"code_sha256_prefix":"7fd1c65224b3375f","entry":"parse_component","repo":"LinesHogan/tLLM","repo_kind":"found_in_text","path":"tllm/common/path_resolution.py","file_url":"https://github.com/LinesHogan/tLLM/blob/HEAD/tllm/common/path_resolution.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":true,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"7fd1c65224b3375f"}},{"code_sha256_prefix":"b28fcc3f898146ac","entry":"pick_common_attn_metadata","repo":"LinesHogan/tLLM","repo_kind":"found_in_text","path":"tllm/common/state.py","file_url":"https://github.com/LinesHogan/tLLM/blob/HEAD/tllm/common/state.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"b28fcc3f898146ac"}},{"code_sha256_prefix":"71b9014bf606bbad","entry":"resolve_object_by_path","repo":"LinesHogan/tLLM","repo_kind":"found_in_text","path":"tllm/common/path_resolution.py","file_url":"https://github.com/LinesHogan/tLLM/blob/HEAD/tllm/common/path_resolution.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"71b9014bf606bbad"}},{"code_sha256_prefix":"070f8e55b4379c68","entry":"resolve_prompt_sample_for_req_id","repo":"LinesHogan/tLLM","repo_kind":"found_in_text","path":"tllm/common/state.py","file_url":"https://github.com/LinesHogan/tLLM/blob/HEAD/tllm/common/state.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"070f8e55b4379c68"}}]},"arxiv_metadata":{"licence":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license)","fields":["title","abstract","authors","date"],"primary_category":"cs.LG","source":"arxiv_2026.jsonl"},"syntology_extracted_results":null}