{"url":"/dataset/timegraph","name":"TimeGraph","full_name":"TimeGraph: Synthetic Benchmark Datasets for Robust Time-Series Causal Discovery","description_markdown":"TimeGraph is a comprehensive suite of synthetic datasets designed to benchmark causal discovery algorithms on time-series data. The dataset captures real-world complexities by incorporating temporal dynamics such as trends, seasonality, and nonstationarity, as well as sampling challenges including irregular time intervals and structured missingness. It features diverse noise types, including Gaussian, heavy-tailed, and heteroskedastic variations, and supports scenarios with latent confounding to enable evaluation under partially observed systems. The underlying causal structures span both linear and nonlinear relationships, including polynomial and trigonometric forms.\r\n\r\nThe motivation behind TimeGraph is to address the current lack of robust and realistic benchmarks in time-series causal discovery. Existing datasets often overlook the intricate challenges observed in real-world domains such as Earth system science, healthcare, and economics.\r\n\r\nTimeGraph serves as a unified testbed for comparing linear and nonlinear causal discovery methods (e.g., PCMCI+, Granger causality, NOTEARS, etc), evaluating algorithmic robustness under conditions of missing data and irregular sampling, training and validating deep causal models and representation learning frameworks, and analyzing the sensitivity of methods to confounding and autocorrelation.","description_withheld":null,"homepage":"https://github.com/hferdous/TimeGraph","introduced_date":"2025-06-02","introduced_date_note":null,"introduced_by":{"paper":"/paper/timegraph-synthetic-benchmark-datasets-for","title":"TimeGraph: Synthetic Benchmark Datasets for Robust Time-Series Causal Discovery","first_author":"Muhammad Hasan Ferdous","url":null},"license":{"name":"CC BY 4.0","url":"https://creativecommons.org/licenses/by/4.0/"},"modalities":[{"name":"Tabular","url":"/datasets/modality/tabular"},{"name":"Time series","url":"/datasets/modality/time-series"}],"tasks":[{"name":"Time Series","url":"/task/time-series-1","datasets_with_task":"/datasets/task/time-series-1"},{"name":"Causal Discovery","url":"/task/causal-discovery","datasets_with_task":"/datasets/task/causal-discovery"}],"languages":[{"name":"English","url":"/datasets/language/english"}],"variants":["TimeGraph"],"data_loaders":[],"num_papers_in_archive":1,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}