{"url":"/dataset/max-60k","name":"MAX-60K","full_name":"Masked Autoencoder for X-ray Fluorescence 60K Dataset","description_markdown":"The dataset for masked autoencoder for X-ray fluorescence (XRF) is a following development after the dataset [(Chao et al., 2022)](https://doi.org/10.1594/PANGAEA.949225). \r\nBesides the published XRF spectra-target measurements (CaCO3 and TOC) pairs of data, we further upload the XRF spectra in that project but without alignments of the target measurements here.\r\nAs the first XRF large dataset compiled in a ML friendly format, we expect to kickoff more ML studies in the field of XRF and geology, especially DL studies. \r\n\r\nThe investigated cores, which form the datast, are mostly retrieved across the high- to mid-latitude Northwest Pacific (37°N-52°N) and the Pacific sector of the Southern Ocean (53°S-63°S), with a water depth coverage from 1211 to 4853 m: <br>\r\n1. Cruise SO264 in the subarctic Northwest Pacific with R/V SONNE in 2018<br>\r\n2. Cruise PS97 in the central Drake Passage with RV Polarstern in 2016<br>\r\n3. Cruise PS75 in the Pacific sector of the Southern Ocean in 2009/2010.<br>\r\n4. Cruise KOMEX I and KOMEX II with R/V Akademik Lavrentyev in 1998 and cruise SO178 in 2004 in the Okhotsk Sea.","description_withheld":null,"homepage":"https://huggingface.co/datasets/paoyw/max-dataset","introduced_date":"2024-10-16","introduced_date_note":null,"introduced_by":{"paper":"/paper/max-masked-autoencoder-for-x-ray-fluorescence","title":"MAX: Masked Autoencoder for X-ray Fluorescence in Geological Investigation","first_author":"An-Sheng Lee","url":null},"license":{"name":"This work is licensed under a CC BY 4.0 license.","url":null},"modalities":[{"name":"Environment","url":"/datasets/modality/environment"}],"tasks":[{"name":"Self-Supervised Learning","url":"/task/self-supervised-learning","datasets_with_task":"/datasets/task/self-supervised-learning"},{"name":"regression","url":"/task/regression-1","datasets_with_task":"/datasets/task/regression-1"},{"name":"Fill Mask","url":"/task/fill-mask","datasets_with_task":"/datasets/task/fill-mask"},{"name":"Chemical Process","url":"/task/chemical-process","datasets_with_task":"/datasets/task/chemical-process"}],"languages":[],"variants":["MAX-60K"],"data_loaders":[],"num_papers_in_archive":1,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}