{"url":"/dataset/wtw","name":"WTW","full_name":"Wired Table in the Wild","description_markdown":"**WTW** (Wired Table in the Wild)  is a large-scale dataset which includes well-annotated structure parsing of multiple style tables in several scenes like the photo, scanning files, web pages.\r\n\r\nWTW dataset has 10970 training samples and 3611 testing ones. The test images are divided into 7 challenging categories.\r\n\r\nDataset for trains and test contain images and labels. The label is in XML format, which has cell bbox and the structure label, includes start row, end row, start col, end col, and table id. In addition, the test set also contains separate file descripts sub-classification information for each image.","description_withheld":null,"homepage":"https://github.com/wangwen-whu/wtw-dataset","introduced_date":"2021-09-06","introduced_date_note":null,"introduced_by":{"paper":"/paper/parsing-table-structures-in-the-wild","title":"Parsing Table Structures in the Wild","first_author":"Rujiao Long","url":null},"license":{"name":"CC BY-NC 4.0","url":"https://github.com/wangwen-whu/WTW-Dataset/blob/main/License"},"modalities":[{"name":"Images","url":"/datasets/modality/images"}],"tasks":[{"name":"Table Recognition","url":"/task/table-recognition","datasets_with_task":"/datasets/task/table-recognition"}],"languages":[],"variants":["WTW"],"data_loaders":[],"num_papers_in_archive":18,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[{"leaderboard":"/sota/table-recognition-on-wtw","task":"Table Recognition","dataset_variant":"WTW","rows":1,"metrics":["F1"],"first_row_in_archive_order":{"model":"StrucTexTv2 (small)","paper":"/paper/structextv2-masked-visual-textual-prediction","metrics":{"F1":"78.9%"},"code_links":[{"title":"PaddlePaddle/VIMER","url":"https://github.com/PaddlePaddle/VIMER/tree/main/StrucTexT/v2"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"}],"papers_with_a_benchmark_row":[{"paper":"/paper/structextv2-masked-visual-textual-prediction","title":"StrucTexTv2: Masked Visual-Textual Prediction for Document Image Pre-training","date":"2023-03-01","rows_on_this_dataset":1,"code_links":1,"syntology":null}],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}