{"url":"/dataset/propbank-pt","name":"PropBank-PT","full_name":null,"description_markdown":"The PropBankPT (Branco et al., 2012) is a set of sentences annotated with their constituency structure and semantic role tags, composed of 3,406 sentences and 44,598 tokens taken from the Wall Street Journal translated.\nFor the creation of this PropBank we adopted a semi-automatic analysis with a double-blind annotation followed by adjudication. The resulting dataset contains three information levels: phrase constituency, grammatical functions, and phrase semantic roles.\nThe main motivation behind the creation of this resource was to build a high quality data set with semantic information that could support the development of automatic semantic role labelers for Portuguese.\nThe development of this resource started under the METANET4U project (at: http://metanet4u.eu/) whose main goal is to contribute to the establishment of a pan-European digital platform that makes available language resources and services, encompassing both datasets and software tools, for speech and language processing, and supports a new generation of exchange facilities for them.\nYou may also be interested in the related resources DeepBankPT, TreeBankPT, DependencyBankPT and LogicalFormBankPT, also available from this repository.","description_withheld":null,"homepage":"https://portulanclarin.net/repository/browse/propbankpt/e69034fa616e11e2a2aa782bcb07413563a2d5c389e04e779bf58a3b09dc588e/","introduced_date":"2013-01-18","introduced_date_note":null,"introduced_by":null,"license":{"name":"CC BY-SA 4.0","url":"https://creativecommons.org/licenses/by-sa/4.0/"},"modalities":[{"name":"Texts","url":"/datasets/modality/texts"}],"tasks":[{"name":"Semantic Role Labeling","url":"/task/semantic-role-labeling","datasets_with_task":"/datasets/task/semantic-role-labeling"}],"languages":[{"name":"Portuguese","url":"/datasets/language/portuguese"}],"variants":["PropBank-PT"],"data_loaders":[],"num_papers_in_archive":0,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}