{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/data-efficiency-and-long-term-prediction","title":"Data efficiency and long term prediction capabilities for neural operator surrogate models of core and edge plasma codes","arxiv_id":"2402.08561","date":"2024-02-13","proceeding":null,"authors":["N. Carey","L. Zanisi","S. Pamela","V. Gopakumar","J. Omotani","J. Buchanan","J. Brandstetter"],"abstract":"Simulation-based plasma scenario development, optimization and control are crucial elements towards the successful deployment of next-generation experimental tokamaks and Fusion power plants. Current simulation codes require extremely intensive use of HPC resources that make them unsuitable for iterative or real time applications. Neural network based surrogate models of expensive simulators have been proposed to speed up such costly workflows. Current efforts in this direction in the Fusion community are mostly limited to point estimates of quantities of interest or simple 1D PDE models, with a few notable exceptions. While the AI literature on methods for neural PDE surrogate models is rich, performance benchmarks for Fusion-relevant 2D fields has so far remained flimited. In this work neural PDE surrogates are trained for the JOREK MHD code and the STORM scrape-off layer code using the PDEArena library (https://github.com/microsoft/pdearena). The performance of these surrogate models is investigated as a function of training set size as well as for long-term predictions. The performance of surrogate models that are trained on either one variable or multiple variables at once is also considered. It is found that surrogates that are trained on more data perform best for both long- and short-term predictions. Additionally, surrogate models trained on multiple variables achieve higher accuracy and more stable performance. Downsampling the training set in time may provide stability in the long term at the expense of the short term predictive capability, but visual inspection of the resulting fields suggests that multiple metrics should be used to evaluate performance.","url_abs":"https://arxiv.org/abs/2402.08561v1","url_pdf":"https://arxiv.org/pdf/2402.08561v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"links_only","authors_date_abstract":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license), from the Kaggle arXiv metadata snapshot of 2026-09-12"},"code_links":[{"paper_slug":"data-efficiency-and-long-term-prediction","repo_url":"https://github.com/microsoft/pdearena","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":0,"framework":"pytorch","reach":{"status":"ok","spdx":"MIT"}}],"tasks":[],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"syntology_url":null,"atlas_url":"https://app.syntology.ai/?focus=2402.08561","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2402.08561"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/microsoft/pdearena","reach":{"status":"ok","spdx":"MIT"}}],"summary":{"ran":2,"ran_draft_wrong":1,"unverified":4},"by_repo_kind":{"official":{"samples":7,"ran":3,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":0,"samples":[{"code_sha256_prefix":"be31a0618c98c1b9","entry":"batchmul1d","repo":"microsoft/pdearena","repo_kind":"official","path":"pdearena/modules/fourier.py","file_url":"https://github.com/microsoft/pdearena/blob/HEAD/pdearena/modules/fourier.py","link_basis":"harvester_set","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"be31a0618c98c1b9"}},{"code_sha256_prefix":"4c0fb383786dadb3","entry":"batchmul2d","repo":"microsoft/pdearena","repo_kind":"official","path":"pdearena/modules/fourier.py","file_url":"https://github.com/microsoft/pdearena/blob/HEAD/pdearena/modules/fourier.py","link_basis":"harvester_set","language":"python","status":"ran_draft_wrong","verification_level":1,"contract_check":"MISDECLARED","metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"4c0fb383786dadb3"}},{"code_sha256_prefix":"d83235be62a2525b","entry":"pearson_correlation","repo":"microsoft/pdearena","repo_kind":"official","path":"pdearena/modules/loss.py","file_url":"https://github.com/microsoft/pdearena/blob/HEAD/pdearena/modules/loss.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"d83235be62a2525b"}},{"code_sha256_prefix":"6c3ea29ff7061656","entry":"batchmul3d","repo":"microsoft/pdearena","repo_kind":"official","path":"pdearena/modules/fourier.py","file_url":"https://github.com/microsoft/pdearena/blob/HEAD/pdearena/modules/fourier.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"6c3ea29ff7061656"}},{"code_sha256_prefix":"86d0930a8092c898","entry":"custommse_loss","repo":"microsoft/pdearena","repo_kind":"official","path":"pdearena/modules/loss.py","file_url":"https://github.com/microsoft/pdearena/blob/HEAD/pdearena/modules/loss.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"86d0930a8092c898"}},{"code_sha256_prefix":"e363d96b6681e2f0","entry":"get_model","repo":"microsoft/pdearena","repo_kind":"official","path":"pdearena/models/pderefiner.py","file_url":"https://github.com/microsoft/pdearena/blob/HEAD/pdearena/models/pderefiner.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"e363d96b6681e2f0"}},{"code_sha256_prefix":"85fc41048fade072","entry":"scaledlp_loss","repo":"microsoft/pdearena","repo_kind":"official","path":"pdearena/modules/loss.py","file_url":"https://github.com/microsoft/pdearena/blob/HEAD/pdearena/modules/loss.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"85fc41048fade072"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}