{"url":"/dataset/synthetic-and-real-apache-log-records","name":"Synthetic and Real Apache Log Records","full_name":"Dataset for the paper \"On Automatic Parsing of Log Records\"","description_markdown":"Each file contains a specific dataset described in the [paper](https://arxiv.org/abs/2102.06320) \"On Automatic Parsing of Log Records\". For example, `T_E.txt` contains the data for the dataset $T_E$. \r\n\r\nIn a file, each log string resides on a separate line and contains a 2-tuple separated by tab (`\\t`). The first element of the tuple is the actual log string that has to be parsed. The second element is the corresponding “translation” specifying the field name for each of the characters of the first element.","description_withheld":null,"homepage":"https://zenodo.org/record/4536514","introduced_date":"2021-02-12","introduced_date_note":null,"introduced_by":{"paper":"/paper/on-automatic-parsing-of-log-records","title":"On Automatic Parsing of Log Records","first_author":"Jared Rand","url":null},"license":{"name":"CC BY 4.0","url":"https://creativecommons.org/licenses/by/4.0/legalcode"},"modalities":[],"tasks":[],"languages":[],"variants":["Synthetic and Real Apache Log Records"],"data_loaders":[],"num_papers_in_archive":1,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}