{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/prompting-in-context-operator-learning-with","title":"Fine-Tune Language Models as Multi-Modal Differential Equation Solvers","arxiv_id":"2308.05061","date":"2023-08-09","proceeding":null,"authors":["Liu Yang","Siting Liu","Stanley J. Osher"],"abstract":"In the growing domain of scientific machine learning, in-context operator learning has shown notable potential in building foundation models, as in this framework the model is trained to learn operators and solve differential equations using prompted data, during the inference stage without weight updates. However, the current model's overdependence on function data overlooks the invaluable human insight into the operator. To address this, we present a transformation of in-context operator learning into a multi-modal paradigm. In particular, we take inspiration from the recent success of large language models, and propose using \"captions\" to integrate human knowledge about the operator, expressed through natural language descriptions and equations. Also, we introduce a novel approach to train a language-model-like architecture, or directly fine-tune existing language models, for in-context operator learning. We beat the baseline on single-modal learning tasks, and also demonstrated the effectiveness of multi-modal learning in enhancing performance and reducing function data requirements. The proposed method not only significantly enhanced the development of the in-context operator learning paradigm, but also created a new path for the application of language models.","url_abs":"https://arxiv.org/abs/2308.05061v4","url_pdf":"https://arxiv.org/pdf/2308.05061v4.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"prompting-in-context-operator-learning-with","repo_url":"https://github.com/liuyangmage/in-context-operator-networks","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":1,"framework":"jax","reach":{"status":"ok","spdx":"MIT"}}],"tasks":[{"task_slug":"efficient-neural-network","task_name":"Efficient Neural Network"},{"task_slug":"language-modeling","task_name":"Language Modeling"},{"task_slug":"language-modelling","task_name":"Language Modelling"},{"task_slug":"operator-learning","task_name":"Operator learning"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2308.05061","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2308.05061"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/liuyangmage/in-context-operator-networks","reach":{"status":"ok","spdx":"MIT"}}],"summary":{"ran_draft_wrong":1,"ran":2,"unverified":1},"by_repo_kind":{"official":{"samples":4,"ran":3,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":0,"samples":[{"code_sha256_prefix":"ea1821fedc537073","entry":"append_dict_list","repo":"liuyangmage/in-context-operator-networks","repo_kind":"official","path":"icon-lm/operator_weno/analysis.py","file_url":"https://github.com/liuyangmage/in-context-operator-networks/blob/HEAD/icon-lm/operator_weno/analysis.py","link_basis":"plan_row","language":"python","status":"ran_draft_wrong","verification_level":1,"contract_check":"OUTPUT_MISDECLARED","metamorphic_tier":"deterministic","behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"ea1821fedc537073"}},{"code_sha256_prefix":"ccc3d008753c19f3","entry":"get_error","repo":"liuyangmage/in-context-operator-networks","repo_kind":"official","path":"icon-lm/operator_weno/analysis_plot.py","file_url":"https://github.com/liuyangmage/in-context-operator-networks/blob/HEAD/icon-lm/operator_weno/analysis_plot.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"ccc3d008753c19f3"}},{"code_sha256_prefix":"53e1fee72ab7a268","entry":"get_error_tune","repo":"liuyangmage/in-context-operator-networks","repo_kind":"official","path":"icon-lm/operator_weno/analysis_plot.py","file_url":"https://github.com/liuyangmage/in-context-operator-networks/blob/HEAD/icon-lm/operator_weno/analysis_plot.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"53e1fee72ab7a268"}},{"code_sha256_prefix":"c4bac98236c1fc2c","entry":"load_tf_weights_in_gpt2","repo":"liuyangmage/in-context-operator-networks","repo_kind":"official","path":"icon-lm/models_gpt2_source.py","file_url":"https://github.com/liuyangmage/in-context-operator-networks/blob/HEAD/icon-lm/models_gpt2_source.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"c4bac98236c1fc2c"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}