Browse State-of-the-Art › Knowledge Base Construction
Knowledge Base Construction
28 papers with code · 0 benchmarks · 0 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
No dataset record in the archive lists this task.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
28 shown of 28 papers with code (94 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
3 Oct 2022 3 repositories listedIn this paper, we present the first corpus of Web tables created specifically out of Russian language material.
-
28 Apr 2019 3 repositories listedIn this paper, we release, describe, and analyze an OIE corpus called OPIEC, which was extracted from the text of English Wikipedia.
-
10 Dec 2022 2 repositories listedThe ability to understand and generate similes is an imperative step to realize human-level AI.
-
7 Jan 2021 2 repositories listedThere has been a steady need to precisely extract structured knowledge from the web (i.
-
7 Nov 2024 1 repository listedHowever, most approaches investigate one question at a time via modest-sized pre-defined samples, introducing an ``availability bias'' (Tversky&Kahnemann, 1973) that prevents the analysis of knowledge (or beliefs) of…
-
27 Aug 2024 1 repository listedWe introduce SHADOW, a fine-tuned language model trained on an intermediate task using associative deductive reasoning, and measure its performance on a knowledge base construction task using Wikidata triple completion.
-
30 Jan 2024 1 repository listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)For each of these CRUD categories, we have developed comprehensive datasets to evaluate the performance of RAG systems.
-
12 Oct 2023 1 repository listedTo address this, we present Vocabulary Expandable BERT for knowledge base construction, which expand the language model's vocabulary while preserving semantic embeddings for newly added words.
-
6 Oct 2022 1 repository listedConstructing a comprehensive, accurate, and useful scientific knowledge base is crucial for human researchers synthesizing scientific knowledge and for enabling Al-driven scientific discovery.
-
26 Aug 2022 1 repository listedOur system is the winner of track 1 of the LM-KBC challenge, based on BERT LM; it achieves 55.
-
23 Aug 2022 1 repository listed Syntology ran 6 of 6 samples · 0 unverifiedProP implements a multi-step approach that combines a variety of prompting techniques to achieve this.
-
29 Jun 2022 1 repository listedWe provide both a human-annotated test dataset and an auto-generated dataset.
-
16 Oct 2021 1 repository listedTo achieve this, it is crucial to represent multilingual knowledge in a shared/unified space.
-
1 Jun 2021 1 repository listedWe use a German WordNet equivalent, GermaNet, to automatically generate training data for German general entity typing.
-
24 Nov 2020 1 repository listedTo take advantage of the high accuracy of human annotation and the cheap cost of distant supervision, we propose the dual supervision framework which effectively utilizes both types of data.
-
7 Oct 2020 1 repository listedPredicting missing facts in a knowledge graph (KG) is a crucial task in knowledge base construction and reasoning, and it has been the subject of much research in recent works using KG embeddings.
-
1 Aug 2020 1 repository listedFor each news source, the annotation starts on random samples of news articles and continues with samples that are drawn using active learning.
-
11 Jun 2020 1 repository listedFrom a corpus of computer science papers on arXiv, we find that our method achieves a Precision@1000 of 99%, compared to 86% for prior work, and a substantially better precision-yield trade-off across the top 15, 000…
-
14 Feb 2020 1 repository listedSince it can be expensive to obtain training data to learn to extract implications for each new domain of reviews, we propose an unsupervised KBC system, Sampo, Specifically, Sampo is tailored to build KBs for domains…
-
9 Nov 2019 1 repository listed Syntology ran 2 of 5 samples · 3 unverified · 5 pointer-only (licence)We use WikiGenderBias to evaluate systems for bias and find that NRE systems exhibit gender biased predictions and lay groundwork for future evaluation of bias in NRE.
-
12 Jun 2019 1 repository listed Syntology ran 2 of 3 samples · 1 unverifiedWe present the first comprehensive study on automatic knowledge base construction for two prevalent commonsense knowledge graphs: ATOMIC (Sap et al., 2019) and ConceptNet (Speer et al., 2017).
-
31 Jul 2018 1 repository listedNowadays, editors tend to separate different subtopics of a long Wiki-pedia article into multiple sub-articles.
-
20 Jan 2018 1 repository listedWeb 2.
-
15 Dec 2016 1 repository listedIn this paper, we describe the construction of TeKnowbase, a knowledge-base of technical concepts in computer science.
-
21 Apr 2016 1 repository listedIn experimental results on the FB15k-237 benchmark we demonstrate that we can match the performance of a comparable model with explicit entity pair representations using a model of attention over relation types.
-
19 Nov 2015 1 repository listedIn response, this paper introduces significant further improvements to the coverage and flexibility of universal schema relation extraction: predictions for entities unseen in training and multilingual transfer learning…
-
26 May 2015 1 repository listedIn this paper, we focus on the three components of a practical system integrating logical and distributional models: 1) Parsing and task representation is the logic-based part where input problems are represented in…
-
1 May 2015 1 repository listed
Syntology lines on 4 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections