{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/arxiv-2608-21386","title":"Model of Models: When Does Emitting a Specialist Beat Attending, Adapting, or Tuning?","arxiv_id":"2608.21386","date":"2026-07-20","proceeding":null,"authors":["John C. Howell"],"abstract":"Given a task described by a few examples, how should a model be specialized to it? Four mechanisms are available -- zero-shot, in-context attention, test-time gradient adaptation, and emitting specialist weights from a hypernetwork -- yet the operating regime of the last is rarely mapped. We run the identical four-way comparison across six tasks spanning regression, generation, language modeling, reinforcement learning, and clinical and genomic classification, holding the specialist, the context, and (where we can) the training budget fixed. The clearest wins for emission are about cost at matched quality: it ties the state-of-the-art amortized tabular model (TabPFN) on clinical few-shot classification while emitting a reusable specialist instead of re-attending the support set per query, and reaches noise-floor shape generation with a $132$-float per-instance program. On few-shot sinusoid regression it is $2$--$3$ orders of magnitude below MAML at zero test-time gradient steps -- a margin that narrows to $\\sim$$30\\times$ but persists once training budgets are equalized. Emission cannot match in-context attention on high-dimensional sequence modeling: under matched-budget pre-training a one-pass adapter recovers only a minority of the in-context gain ($14.0\\pm0.9\\%$ at $5$M, $11.2\\pm0.5\\%$ at $15$M), and a LoRA-rank sweep shows this shortfall is a partial capacity limit -- capture climbs from $5\\%$ to $21\\%$ as rank grows but plateaus far below full recovery. Mechanism ablations confirm the emitted specialist is genuinely task-conditioned, not a memorized prior; and, more speculatively, emitted specialists compose in weight space -- interpolating two of them tracks the corresponding blend of their functions. We close with a falsifiable thesis, operationalized through a per-task resolution measure, bounding when each conditioning mechanism should be preferred.","url_abs":"https://arxiv.org/abs/2608.21386","url_pdf":"https://arxiv.org/pdf/2608.21386","source":{"archive":null,"snapshot":"2025-07-28","note":"not in the Papers with Code archive (frozen at the snapshot)","row_kind":"graph","title_abstract_authors_date":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license)"},"code_links":[],"tasks":[],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":null,"mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2608.21386"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"mentioned_in_github":null,"is_official":null,"provenance":"deterministic:regex_extraction","mentioned_in_paper":null,"url":"https://github.com/johnchowell/Model-of-Models","reach":null}],"summary":{"ran":6,"unverified":3},"by_repo_kind":{"found_in_text":{"samples":9,"ran":6,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":0,"samples":[{"code_sha256_prefix":"dfbbaf8bada6e08d","entry":"chunk_bpc","repo":"johnchowell/Model-of-Models","repo_kind":"found_in_text","path":"exp13_hyperlora.py","file_url":"https://github.com/johnchowell/Model-of-Models/blob/HEAD/exp13_hyperlora.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"dfbbaf8bada6e08d"}},{"code_sha256_prefix":"8b3aae2b84b3210e","entry":"exact_match","repo":"johnchowell/Model-of-Models","repo_kind":"found_in_text","path":"exp16_arc.py","file_url":"https://github.com/johnchowell/Model-of-Models/blob/HEAD/exp16_arc.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":true,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"8b3aae2b84b3210e"}},{"code_sha256_prefix":"177e6731706681d8","entry":"load_corpus","repo":"johnchowell/Model-of-Models","repo_kind":"found_in_text","path":"exp12_lm.py","file_url":"https://github.com/johnchowell/Model-of-Models/blob/HEAD/exp12_lm.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"177e6731706681d8"}},{"code_sha256_prefix":"a7d7bb2a40592c1f","entry":"lora_delta","repo":"johnchowell/Model-of-Models","repo_kind":"found_in_text","path":"exp12_lm.py","file_url":"https://github.com/johnchowell/Model-of-Models/blob/HEAD/exp12_lm.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"a7d7bb2a40592c1f"}},{"code_sha256_prefix":"41ee7c197b144bef","entry":"sample_lm_batch","repo":"johnchowell/Model-of-Models","repo_kind":"found_in_text","path":"exp12_lm.py","file_url":"https://github.com/johnchowell/Model-of-Models/blob/HEAD/exp12_lm.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"41ee7c197b144bef"}},{"code_sha256_prefix":"6835abbeb38e0957","entry":"sample_tasks","repo":"johnchowell/Model-of-Models","repo_kind":"found_in_text","path":"exp14_sinusoid.py","file_url":"https://github.com/johnchowell/Model-of-Models/blob/HEAD/exp14_sinusoid.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"6835abbeb38e0957"}},{"code_sha256_prefix":"36e28ee9fa2ee1a5","entry":"evaluate_real","repo":"johnchowell/Model-of-Models","repo_kind":"found_in_text","path":"exp16_arc.py","file_url":"https://github.com/johnchowell/Model-of-Models/blob/HEAD/exp16_arc.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"36e28ee9fa2ee1a5"}},{"code_sha256_prefix":"9a1ed389e8887d97","entry":"fourier","repo":"johnchowell/Model-of-Models","repo_kind":"found_in_text","path":"exp11_hypernet.py","file_url":"https://github.com/johnchowell/Model-of-Models/blob/HEAD/exp11_hypernet.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"9a1ed389e8887d97"}},{"code_sha256_prefix":"2d88b7ef3bc824ca","entry":"greedy_step","repo":"johnchowell/Model-of-Models","repo_kind":"found_in_text","path":"exp10_ablate.py","file_url":"https://github.com/johnchowell/Model-of-Models/blob/HEAD/exp10_ablate.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"2d88b7ef3bc824ca"}}]},"arxiv_metadata":{"licence":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license)","fields":["title","abstract","authors","date"],"primary_category":"cs.LG","source":"arxiv_2026.jsonl"},"syntology_extracted_results":null}