{"url":"/dataset/open-pi","name":"Open PI","full_name":"Open PI","description_markdown":"**Open PI** is the first dataset for tracking state changes in procedural text from arbitrary domains by using an unrestricted (open) vocabulary. The dataset comprises 29,928 state changes over 4,050 sentences from 810 procedural real-world paragraphs from WikiHow.com.\r\nThe state tracking task assumes new formulation in which just the text is provided, from which a set of state changes (entity, attribute, before, after) is generated for each step, where the entity, attribute, and values must all be predicted from an open vocabulary.\r\n\r\nSource: [Allen Institute of AI](https://allenai.org/data/openpi)","description_withheld":null,"homepage":"https://allenai.org/data/openpi","introduced_date":"2020-10-31","introduced_date_note":null,"introduced_by":{"paper":"/paper/a-dataset-for-tracking-entities-in-open","title":"A Dataset for Tracking Entities in Open Domain Procedural Text","first_author":"Niket Tandon","url":null},"license":null,"modalities":[{"name":"Texts","url":"/datasets/modality/texts"}],"tasks":[],"languages":[],"variants":["Open PI"],"data_loaders":[],"num_papers_in_archive":11,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}