{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/rapid-word-learning-through-meta-in-context","title":"Rapid Word Learning Through Meta In-Context Learning","arxiv_id":"2502.14791","date":"2025-02-20","proceeding":null,"authors":["Wentao Wang","Guangyuan Jiang","Tal Linzen","Brenden M. Lake"],"abstract":"Humans can quickly learn a new word from a few illustrative examples, and then systematically and flexibly use it in novel contexts. Yet the abilities of current language models for few-shot word learning, and methods for improving these abilities, are underexplored. In this study, we introduce a novel method, Meta-training for IN-context learNing Of Words (Minnow). This method trains language models to generate new examples of a word's usage given a few in-context examples, using a special placeholder token to represent the new word. This training is repeated on many new words to develop a general word-learning ability. We find that training models from scratch with Minnow on human-scale child-directed language enables strong few-shot word learning, comparable to a large language model (LLM) pre-trained on orders of magnitude more data. Furthermore, through discriminative and generative evaluations, we demonstrate that finetuning pre-trained LLMs with Minnow improves their ability to discriminate between new words, identify syntactic categories of new words, and generate reasonable new usages and definitions for new words, based on one or a few in-context examples. These findings highlight the data efficiency of Minnow and its potential to improve language model performance in word learning tasks.","url_abs":"https://arxiv.org/abs/2502.14791v2","url_pdf":"https://arxiv.org/pdf/2502.14791v2.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[],"tasks":[{"task_slug":"in-context-learning","task_name":"In-Context Learning"},{"task_slug":"language-modeling","task_name":"Language Modeling"},{"task_slug":"language-modelling","task_name":"Language Modelling"},{"task_slug":"large-language-model","task_name":"Large Language Model"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"syntology_url":"https://syntology.ai/paper/2502.14791","atlas_url":"https://app.syntology.ai/?focus=2502.14791","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2502.14791"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-25T09:33:49+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"deterministic:regex_extraction","url":"https://github.com/wwt17/meta-learning-word","reach":null}],"summary":{"ran":1,"ran_fixture":2,"ran_draft_wrong":1},"by_repo_kind":{"found_in_text":{"samples":4,"ran":4,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":4,"samples":[{"code_sha256_prefix":"265c98fea0521afa","entry":"InContextFormat","repo":"wwt17/meta-learning-word","repo_kind":"found_in_text","path":"in_context_format.py","file_url":"https://github.com/wwt17/meta-learning-word/blob/HEAD/in_context_format.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"265c98fea0521afa"}},{"code_sha256_prefix":"5d38f913beb049c9","entry":"_offset_with_leading_space","repo":"wwt17/meta-learning-word","repo_kind":"found_in_text","path":"in_context_format.py","file_url":"https://github.com/wwt17/meta-learning-word/blob/HEAD/in_context_format.py","link_basis":"first_harvest_node","language":"python","status":"ran_fixture","verification_level":1,"contract_check":"RAISES","metamorphic_tier":"well_formed","behaviour_fingerprint":true,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"5d38f913beb049c9"}},{"code_sha256_prefix":"fcfa8ce0043c26f1","entry":"example_str","repo":"wwt17/meta-learning-word","repo_kind":"found_in_text","path":"in_context_format.py","file_url":"https://github.com/wwt17/meta-learning-word/blob/HEAD/in_context_format.py","link_basis":"first_harvest_node","language":"python","status":"ran_draft_wrong","verification_level":1,"contract_check":"OUTPUT_MISDECLARED","metamorphic_tier":"deterministic","behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"fcfa8ce0043c26f1"}},{"code_sha256_prefix":"42cf29c8b71df1c4","entry":"replace_at_offsets","repo":"wwt17/meta-learning-word","repo_kind":"found_in_text","path":"in_context_format.py","file_url":"https://github.com/wwt17/meta-learning-word/blob/HEAD/in_context_format.py","link_basis":"first_harvest_node","language":"python","status":"ran_fixture","verification_level":1,"contract_check":"RAISES","metamorphic_tier":"deterministic","behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"42cf29c8b71df1c4"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}