{"url":"/dataset/fakbat","name":"FAKBAT","full_name":null,"description_markdown":"The Freebase Annotations of TREC KBA 2014 Stream Corpus with Timestamps (**FAKBAT**) is an extension of the FAKBA1 dataset that contains entity age and entity timestamp. It comprises roughly 1.2 billion timestamped documents from global public news wires, blogs, forums, and shortened links shared on social media. It spans 572 days (October 7, 2011–May 1, 2013).\n\nSource: [https://arxiv.org/pdf/1701.04039.pdf](https://arxiv.org/pdf/1701.04039.pdf)","description_withheld":null,"homepage":"https://github.com/graus/emerging-entities-timeseries","introduced_date":null,"introduced_date_note":null,"introduced_by":{"paper":"/paper/the-birth-of-collective-memories-analyzing","title":"The Birth of Collective Memories: Analyzing Emerging Entities in Text Streams","first_author":"David Graus","url":null},"license":null,"modalities":[{"name":"Texts","url":"/datasets/modality/texts"}],"tasks":[],"languages":[],"variants":["FAKBAT"],"data_loaders":[{"repo":"https://github.com/graus/emerging-entities-timeseries","url":"https://github.com/graus/emerging-entities-timeseries","frameworks":[]}],"num_papers_in_archive":2,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}